
Google发布了三个新的Gemini模型,重点是“输出更少的token,成本更低”,在候诊室测试Gemini 4 - 并让Gemini 3.5 Pro留在起跑线上。
用简单的语言来说 - Google于2026年7月21日发布了三个新的Gemini变体,以Gemini 3.6 Flash为主打。其核心卖点很简单:输出令牌更少,价格更低,并且混合了一个针对网络安全调优的模型。但Gemini 3.5 Pro——这个被预告了数周的旗舰推理模型——仍在测试中。Google反而提前宣布了Gemini 4。
Google在2026年的Gemini节奏并不顺利:3.5 Pro多次延期,Google私下称编码基准回归。发布“现有就绪”产品(Flash及相关专用变体)的方法是一种在推理旗舰重新调整的同时保持发布节奏的方式。
正在发生两件事。首先,Google正在致力于双层模型现实:Flash类模型优化成本和延迟,Pro/Ultra类用于最困难的推理。这与OpenAI的GPT-5.6与Codex/Work分离以及Anthropic的Fable与Haiku分离相呼应。三大前沿实验室正在收敛到相同的产品拓扑。
其次,在3.5 Pro尚未发布的情况下预告Gemini 4是一个信号,而非泄露。这是Google在向市场和企业客户传达:推理天花板仍在提升,不要锁定在竞争对手上。打赌的是“还有更多即将到来”足以保持管道温暖,即使旗舰产品延迟。
对于采购方的工程师:3.6 Flash的真正价值是可衡量的——输出令牌的减少在生产中会累积。将其与您最差的延迟任务进行测试。对于与CFO的路线图对话:在3.5 Pro实际上GA之前,不要签订多年Gemini合同。
对于建设者:不要等待——3.6 Flash值得现在进行基准测试。对于决策者:双层现实已经到来;购买您实际需要的层级,而不是您想要提及的旗舰产品。
本文由人工智能撰写,并经人工编辑审核。
I hope Gemini 3.6 Flash will handle creative tasks well. Shorter outputs might limit artistic expression.
I wonder if Gemini 3.6 Flash will be suitable for detailed, nuanced discussions. Shorter outputs might not capture the depth needed for complex topics.
I'm interested in seeing how Gemini 4 will compare to the Flash and Pro models. Will it offer a balanced mix of cost and quality?
Gemini 4 might focus on advanced features rather than cost, setting it apart from Flash and Pro.
I'm concerned about the potential lack of depth in Gemini 3.6 Flash. Will it sacrifice quality for brevity?
I wonder how the shorter outputs will impact complex queries. Will Gemini 3.6 Flash still deliver the depth needed for detailed analysis?
I'm curious about the balance between cost and quality in these new models. Will the shorter outputs still provide meaningful insights?
Google's new Gemini models sound promising, but I wonder how the reduced token output will affect the quality of responses.
Fatigue hype 2026 : le tri entre modèle et harness