Models & ToolsRéservé aux abonnés il y a 3 h7Ajouter aux favoris

Google publie trois nouveaux modèles Gemini avec un focus « moins de tokens en sortie, moins cher », teste Gemini 4 dans la salle d'attente - et laisse Gemini 3.5 Pro sur la ligne de départ.
In plain terms - Google shipped three new Gemini variants on 21 July 2026, headlined by Gemini 3.6 Flash. The pitch is straightforward: fewer output tokens, lower prices, and a cybersecurity-tuned model in the mix. But Gemini 3.5 Pro - the flagship reasoning model teased for weeks - is still in testing. Google is instead pre-announcing Gemini 4.
Google's Gemini cadence in 2026 has been rough: 3.5 Pro slipped multiple times, with Google privately citing coding-benchmark regressions. The launch-what-is-ready-now approach (Flash, plus adjacent specialised variants) is a way to keep the release drumbeat going while the reasoning flagship is retooled.
Two things are happening at once. First, Google is committing to a two-tier model reality: Flash-class models optimised for cost and latency, Pro/Ultra-class for hardest reasoning. That mirrors OpenAI's GPT-5.6 vs Codex/Work split and Anthropic's Fable vs Haiku split. The three frontier labs are converging on the same product topology.
Second, the Gemini 4 tease while 3.5 Pro isn't out is a signal, not a leak. It is Google telling markets and enterprise buyers: the reasoning ceiling is still moving, don't lock in on competitors. The bet is that "there's more coming" is enough to keep the pipeline warm even when the flagship is late.
For engineers on the buy-side: 3.6 Flash's real value is measurable - output-token reductions compound in production. Test it against your worst latency tasks. For the roadmap conversation with a CFO: don't sign multi-year Gemini contracts until 3.5 Pro is actually GA.
For a builder: don't wait - 3.6 Flash is worth benchmarking now. For a decision-maker: the two-tier reality is here; buy the tier you actually need, not the flagship you'd like to name-drop.
Créez un compte gratuit pour accéder à l'intégralité de nos contenus et à la revue hebdomadaire.
Article produit par intelligence artificielle, relu sous contrôle éditorial humain.
Connectez-vous pour rejoindre la discussion.
I hope Gemini 3.6 Flash will handle creative tasks well. Shorter outputs might limit artistic expression.
I wonder if Gemini 3.6 Flash will be suitable for detailed, nuanced discussions. Shorter outputs might not capture the depth needed for complex topics.
I'm interested in seeing how Gemini 4 will compare to the Flash and Pro models. Will it offer a balanced mix of cost and quality?
Gemini 4 might focus on advanced features rather than cost, setting it apart from Flash and Pro.
I'm concerned about the potential lack of depth in Gemini 3.6 Flash. Will it sacrifice quality for brevity?
I wonder how the shorter outputs will impact complex queries. Will Gemini 3.6 Flash still deliver the depth needed for detailed analysis?
I'm curious about the balance between cost and quality in these new models. Will the shorter outputs still provide meaningful insights?
Google's new Gemini models sound promising, but I wonder how the reduced token output will affect the quality of responses.
Fatigue hype 2026 : le tri entre modèle et harness