Models & ToolsSubscribers only 41 min ago2Add to bookmarks

Three architectures to dissect side-by-side: Poolside's MoE 118B / 8B active on a single machine; Inkling's multimodal 975B / 41B, also released as 276B / 12B; and Kimi K3's 2.8T / 104B context with 1M tokens, whose commercial agreement clause could exclude US companies.
In plain terms. Three major open models released simultaneously: Laguna S2.1 (Poolside), Inkling (Thinking Machines), and Kimi K3 (Moonshot). They target different use cases—simple deployment, multimodal fine-tuning base, and maximal frontier performance—and their licenses make them distinct strategic choices for a CTO.
Interconnects recap #23 (Nathan Lambert, August 2, 2026) compiles the open week. Three releases stand out, each with a distinct architecture and a licensing agreement that matters as much as the benchmarks.
Segmentation by deployment constraints. Laguna for teams wanting minimal infra (the only one of the three executable on a single documented machine). Inkling 276B / 12B as a lightweight multimodal R&D base. K3 for maximal frontier—but the commercial agreement clause may block a US company under export constraints.
Third-party benchmarks; first K3 implementations outside China; Inkling 276B / 12B version on Hugging Face; effective cost per served token.
So what. This week, open isn’t about one model: it’s three distinct strategic choices, to be sorted by deployment constraints and licensing terms.
Create a free account to access all our content and the weekly review.
Article produced by artificial intelligence, reviewed under human editorial control.
Sign in to join the discussion.
Inkling's multimodal jump is impressive, but I wonder if the 276B/12B variant will be usable on consumer hardware or if that's reserved for enterprise only.
The 118B MoE on Spark is wild-wonder how much latency jumps when you scale to 10 users on one machine.
Kimi K3 : de la preview au live