Black Forest Labs opens early access for multimodal FLUX 3
It integrates image, video, audio, and action-prediction capabilities.

Introducing FLUX 3. One multi-modal model for Image, Video, Audio and Action-Prediction. Creations are truer to life in every kind of style. FLUX 3 Video is now available in early access (link below). Jointly trained in one unified architecture, our model can be extended to predict actions for robotics. See our work with mimic and Audi in the thread.
Combined views
182.5K
19 posts, first seen 6h ago