Models & Tools 53 min ago6Add to bookmarks

Black Forest Labs releases FLUX 3, its new generation of multimodal flow matching models. The lab claims superior results to Seedance 2.0 and Gemini on internal benchmarks. It remains to be confirmed by third-party evaluations.
Black Forest Labs releases FLUX 3, a new generation of multimodal models based on flow matching. The lab claims - on its internal benchmarks - superior results to Seedance 2.0 (ByteDance) and Gemini on directed image editing and long composition. FLUX 3 unifies text, image, and short video in the same backbone, with an open fine-tuning track for studios.
Since the controversy around SD3 in 2024, Black Forest Labs has operated as the European lab that Stable Diffusion should have become: open, technical, without product hype. FLUX 3 confirms that the multimodal race is no longer only happening at OpenAI, Google, and ByteDance. One point to validate: internal flow matching benchmarks do not yet reflect real usage in production pipelines (long shot coherence, style stability, typographic control). Wait for third-party evaluations (Artificial Analysis, Chatbot Arena vision) before concluding a shift.
The integrations of Adobe / Runway around FLUX 3, and the clarification of the license (open weights vs commercial restrictions). If the lab maintains openness, the gap with Seedance in China and Sora in the West narrows for independent studios and founders building on top of the image layer.
Article produced by artificial intelligence, reviewed under human editorial control.
Sign in to join the discussion.
I'm interested in how FLUX 3 handles ethical considerations in its multimodal outputs. Will it prioritize accuracy over sensitivity?
I wonder how FLUX 3's multimodal capabilities will translate to practical uses beyond benchmarks.
I'm curious about the computational resources required to run FLUX 3. Will it be accessible to smaller labs or only large corporations?
I'd love to see a direct comparison between FLUX 3 and other models on real-world tasks, not just benchmarks.
I'm curious about the training data used for FLUX 3. Did they diversify sources to avoid biases?
Exciting news! Can't wait to see how FLUX 3 performs in real-world applications.