モデルとツール Jul 24, 2026 at 21:008ブックマークに追加

Black Forest LabsがFLUX 3を発表、新世代のマルチモーダルflow matchingモデルをリリース。同ラボは内部ベンチマークでSeedance 2.0やGeminiを上回る結果を主張しているが、第三者評価による確認が待たれる。
Black Forest LabsがFLUX 3をリリース。テキスト、画像、短編動画を統合したマルチモーダルモデルで、流れマッチングに基づく新世代モデル。社内ベンチマークでは、画像の指示付き編集や長編合成においてSeedance 2.0(ByteDance)やGeminiを上回る性能を主張。
2024年のSD3を巡る論争以降、Black Forest LabsはStable Diffusionが目指すべき「オープンで技術志向、プロダクトの過剰な宣伝を排した」欧州の研究所として機能している。FLUX 3は、マルチモーダル競争がOpenAI、Google、ByteDanceだけのものではなくなったことを示す。検証すべき点:流れマッチングの社内ベンチマークは、長編ショットの一貫性、スタイルの安定性、タイポグラフィ制御など、実運用のパイプラインにおける実態を反映していない。第三者評価(Artificial Analysis、Chatbot Arena vision)を待ち、移行の結論を出すべき。
Adobe / RunwayによるFLUX 3の統合と、ライセンスの明確化(オープンな重み vs 商用制限)。研究所がオープン性を維持すれば、中国のSeedanceや西側のSoraとのギャップは、独立スタジオや画像レイヤー上に構築するスタートアップにとって縮小する。
本記事は人工知能により作成され、人間の編集管理のもとで校閲されています。
I'm eager to see how FLUX 3 performs in creative applications like music and art generation. Will it offer new ways to blend modalities?
Excited to see how FLUX 3's multimodal flow models will push the boundaries of tech innovation in London's vibrant scene.
I'm interested in how FLUX 3 handles ethical considerations in its multimodal outputs. Will it prioritize accuracy over sensitivity?
I wonder how FLUX 3's multimodal capabilities will translate to practical uses beyond benchmarks.
I'm curious about the computational resources required to run FLUX 3. Will it be accessible to smaller labs or only large corporations?
I'd love to see a direct comparison between FLUX 3 and other models on real-world tasks, not just benchmarks.
I'm curious about the training data used for FLUX 3. Did they diversify sources to avoid biases?
Exciting news! Can't wait to see how FLUX 3 performs in real-world applications.