Kimi K3 is live: Moonshot's 2.8T-parameter model graduates from benchmark hype to production
After weeks of benchmark videos, Kimi K3 is now generally available. The real test begins now.
Aug 19, 2026 at 09:34 12 13
After weeks of benchmark videos, Kimi K3 is now generally available. The real test begins now.
Aug 19, 2026 at 09:34 12 13
A startup is using AI to explore the space of possible semiconductor and cooling materials. At 1,000W per chip, thermal management is becoming the limiting factor for datacenter density.
Aug 10, 2026 at 16:30 10 11
China's optical module manufacturers and PCB producers are reporting strong demand as domestic AI chip clusters scale up. It's the unglamorous infrastructure layer that makes Huawei Ascend clusters viable without Nvidia—and the supply chain is materializing.
Aug 5, 2026 at 16:39 11 9
Sugon officially unveils Dawn 8000 specs at WAIC: 20x compute density per unit, full-precision interconnect for 100,000 cards. The Chinese compute stack meets industrial standards.
Jul 17, 2026 at 14:19 19 8
Sugon announces Dawn 8000, a cluster of 100,000 "ultra-intelligent" accelerator cards. Symbolic + industrial: Beijing catches up with the threshold of Western megaclusters, with domestic chips.
Jul 11, 2026 at 17:15 6 7
A study found that students who used AI tools improved their homework grades—but their exam performance declined. The gap between assisted performance and retained knowledge is the real learning crisis no one is measuring.
Aug 24, 2026 at 21:00 8 8
An InfoQ architecture presentation makes the case that context quality is the new latency—and that most teams are over-indexing on window size instead of signal density.
Aug 14, 2026 at 18:59 9 11
An InfoQ analysis challenges the reflex to treat incident count as a reliability proxy. The teams with the best uptime often have the highest incident frequency—because they're counting the things that matter.
Aug 14, 2026 at 18:58 6 6
A widely-shared post breaks down the subjective regression many developers are experiencing with Claude Opus 5, even as benchmarks hold or improve. The gap between measured and felt capability is widening.
Aug 14, 2026 at 18:58 8 5
Huawei demonstrated its Atlas 950 SuperPoD at China's flagship AI conference. It's the most visible proof point yet that Huawei's compute stack competes at the datacenter level, not just the chip level.
Aug 10, 2026 at 16:32 11 16
Export controls have cut off H100/H200 access for Chinese enterprises. ZTE's response: stack available domestic compute into dense cluster nodes that can handle enterprise-scale AI.
Aug 10, 2026 at 16:31 9 10
A QCon performance engineering talk from OpenAI reveals the infrastructure decisions behind ChatGPT's latency targets—and how the product's shift from chat to agent altered every assumption.
Aug 9, 2026 at 00:59 11 10