Fetching from the wire…
Public story · 2026-07-21 · high
Moonshot's 2.8 trillion parameter model tops Claude on a blind coding eval the same day Bessent floats sanctions on Chinese AI.
Why now: Full weights release July 27, and the sanctions talk broke the same day as the benchmark news, per Bessent's July 21 comments to Fox Business.
Moonshot AI released Kimi K3, a sparse mixture-of-experts model activating just 16 of 896 experts per token, about 1.8% of the pool live at any moment. It has a 1M-token context window, native vision, and two new architectural pieces called Kimi Delta Attention and Attention Residuals. Full weights land July 27 under a modified-MIT-style license, with hosted pricing at $0.30/M cache-hit input, $3/M cache-miss, $15/M output.
Here's the part that should actually change what you build with. Kimi K3 ranked first on Arena's Frontend Code evaluation at 1,679 points, ahead of Claude Fable 5 in blind developer testing. On overall capability it still trails Fable 5 and GPT-5.6 Sol. So this isn't China catching up across the board, it's narrower than that: on the specific job of writing frontend code, an open model with published weights beat the model I use every day in my personal projects.
That news landed the same day Treasury Secretary Scott Bessent told Fox Business the administration will investigate whether Chinese models were distilled from US systems, saying "we are finding watermarks of our U.S. large language models on many of the Chinese models, and that's unacceptable," with sanctions on the table. This follows Anthropic's June 10 letter to Senate Banking alleging Alibaba's Qwen lab ran 28.8 million Claude exchanges through roughly 25,000 fraudulent accounts between April 22 and June 5.
MIT Technology Review reports the White House AI apparatus is split: one camp wants export controls on cheap capable Chinese weights, another thinks the bigger risk is the US retreating from open models. CAISI lost its director three months in (Chris Fall resigned July 20, NIST's Arvind Raman is acting), and a voluntary 30-day classified pre-release review framework with OpenAI, Anthropic, and Google is still being finalized.
I'm not a lawyer. But weights on your own disk survive a policy fight that a hosted API subscription doesn't.
Each link below shares sources, entities, or timing with this story.
The open-weight race just changed constraint. Moonshot AI suspended all new consumer subscriptions on July 20, roughly 48 hours after Kimi K3 launched, because request volume pushed its compute cluster to capacity. Remaining GPUs are reserved for existing paid subscribers. Tec...
Xiaomi released MiMo-V2.5-Pro, a 1.02 trillion parameter mixture-of-experts model (42B active) with 1M token context, fully MIT licensed. In benchmarks, it achieves 63.8% success on agentic tasks using 40-60% fewer tokens than Claude Opus 4.6 or GPT-5.4 for comparable results....
25,000 fake accounts. 28.8 million Claude conversations. Six weeks. And the thing they were harvesting wasn't trivia, it was software engineering and agentic reasoning. In a June 24 letter to US senators and the White House, Anthropic alleged that operators tied to Alibaba's Q...
This is a supply-chain fact, and most people are still treating it as a geopolitics argument. Sequoia published "America's Open-Model Paradox" on July 24 with the number that reframes the whole conversation: Qwen's share of open-model fine-tunes went from 1% in January 2024 to...
Martin Alderson's essay "The upcoming AI margin collapse, part 1: GLM 5.2" hit 675 points and 462 comments on Hacker News, and it's the rare HN chart-topper that's actually about spreadsheet math instead of vibes. The argument is simple. Z.ai's GLM 5.2 delivers frontier-adjace...
GLM-5.1 from Zhipu AI scored 58.4% on SWE-bench Pro. GPT-5.4 scored 57.7%. Claude Opus 4.6 scored 57.3%. That's the first time an open-weight model has ever topped a major coding benchmark against the best proprietary models. The specs matter. GLM-5.1 is a 754B-parameter mixtu...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.