FLUX 3
Overview
FLUX 3 is a multimodal foundation model from Black Forest Labs that jointly learns from images, video, and audio within one unified architecture. It generates and edits video with native synchronized audio up to 20 seconds, supports image and video conditioning, keyframe control, and multilingual dialogue, and produces a wide range of visual styles. Image synthesis and action-prediction capabilities are being rolled out in later early access phases.
About Black Forest Labs
Born from foundational research, we continually build state-of-the-art technology that advances how the world is seen and understood.
View Company Profile