Kimi K3 is live: Moonshot's 2.8T-parameter model graduates from benchmark hype to production
After weeks of benchmark videos, Kimi K3 is generally available. Now the real test begins.
19/08/2026 à 09h34 11 11
After weeks of benchmark videos, Kimi K3 is generally available. Now the real test begins.
19/08/2026 à 09h34 11 11
WeLM, the model powering WeChat's AI agent, has quietly scaled to 617 billion parameters with an undisclosed decoding mechanism. No benchmarks, no papers - just a billion-user deployment.
13/08/2026 à 12h57 9 10
OpenAI has pushed GPT-5.6 Sol into ChatGPT, continuing the strategy of building a tiered capability ladder within a single product - and pressuring competitors at every price point.
07/08/2026 à 10h54 8 10
Une entreprise chinoise, Wiz, publie un modèle IA de prévision au niveau ville : trafic, énergie, air. Angle « planification » plus qu'usage grand public - signe de maturité d'un segment sous-radar.
12/07/2026 à 11h15 7 10
The fintech company Ramp has launched an internal AI model routing layer - and open-sourced it as a product. The move reflects a broader pattern: companies that ship AI at scale are building model management infrastructure that the market hasn't provided.
20/08/2026 à 22h32 10 11
GLM-5.3 from Chinese lab Zhipu AI targets the coding and cybersecurity verticals specifically, claiming 2× improvement on SWE-Marathon benchmarks and topping CyberGym evals. The vertical specialization race is accelerating.
14/08/2026 à 18h58 6 7
A widely-shared post breaks down the subjective regression many developers are experiencing with Claude Opus 5, even as benchmarks hold or improve. The gap between measured and felt capability is widening.
14/08/2026 à 18h58 6 5
Meta's open-weight play just got concrete: a 30B model that runs locally, handles multi-step agent tasks, and doesn't route your code through third-party APIs.
10/08/2026 à 16h29 11 13
Trois architectures à disséquer côte-à-côte : le MoE 118B / 8B actifs de Poolside exécutable sur une seule machine ; le multimodal 975B / 41B d'Inkling, publié aussi en 276B / 12B ; le 2,8T / 104B contexte 1M de Kimi K3, dont la clause d'accord commercial pourrait éjecter les entreprises US.
03/08/2026 à 00h53 16 13
Depuis aujourd'hui, transparence de l'entraînement, divulgation du sourcing protégé et gestion des risques systémiques ne sont plus des recommandations mais des obligations légales pour les modèles à usage général. Reste à savoir si l'AI Office, à peine constitué, saura convertir ce mandat en jurisprudence - sous la menace explicite d'une riposte américaine.
03/08/2026 à 00h52 14 12
Sur le benchmark IPI, Opus 5 divise par près de trois le taux de succès attaquant d'Opus 4.8. Le meilleur non-Claude évalué reste à 16,5 %. Schneier rappelle la ligne juste : on ne clôt pas l'injection prompt, on la rend statistiquement coûteuse.
31/07/2026 à 22h20 12 11
MiniMax publie H3, modèle open-source full-modal (texte, image, audio, vidéo) au tiers du prix des concurrents. Le marché vidéo IA cesse d'être binaire fermé cher / open faible.
31/07/2026 à 12h50 12 14