
Black Forest Labs 推出了其新一代多模态流匹配模型FLUX 3。该实验室声称在内部基准测试中超越了Seedance 2.0和Gemini的成果。仍需通过第三方评估来确认。
Black Forest Labs 发布了基于流匹配的新一代多模态模型 FLUX 3。该实验室在其内部基准测试中声称,在图像指导编辑和长篇组合方面的结果优于 Seedance 2.0(字节跳动)和 Gemini。FLUX 3 在同一主干中统一了文本、图像和短视频,并为工作室提供了一个开放的微调轨道。
自 2024 年 SD3 争议以来,Black Forest Labs 以 Stable Diffusion 本应成为的欧洲实验室的身份运作:开放、技术、无产品炒作。FLUX 3 确认了多模态竞赛不再仅限于 OpenAI、Google 和字节跳动。需要验证的一点是:内部流匹配基准测试尚未反映生产流水线中的实际使用情况(长镜头一致性、风格稳定性、排版控制)。在得出结论之前,请等待第三方评估(Artificial Analysis、Chatbot Arena vision)。
围绕 FLUX 3 的 Adobe / Runway 整合,以及许可证的明确(开放权重与商业限制)。如果该实验室坚持开放,那么与中国的 Seedance 和西方的 Sora 之间的差距将缩小,这对于独立工作室和在图像层之上构建的创始人来说是一个好消息。
本文由人工智能撰写,并经人工编辑审核。
I'm interested in how FLUX 3 handles ethical considerations in its multimodal outputs. Will it prioritize accuracy over sensitivity?
I wonder how FLUX 3's multimodal capabilities will translate to practical uses beyond benchmarks.
I'm curious about the computational resources required to run FLUX 3. Will it be accessible to smaller labs or only large corporations?
I'd love to see a direct comparison between FLUX 3 and other models on real-world tasks, not just benchmarks.
I'm curious about the training data used for FLUX 3. Did they diversify sources to avoid biases?
Exciting news! Can't wait to see how FLUX 3 performs in real-world applications.