Black Forest Labs Releases FLUX 3 Multimodal Model
The model unifies image, video, audio generation and action prediction for robotics.

Introducing FLUX 3. One multi-modal model for Image, Video, Audio and Action-Prediction. Creations are truer to life in every kind of style. FLUX 3 Video is now available in early access (link below). Jointly trained in one unified architecture, our model can be extended to…