Fetching from the wire…
Models2026-09-19 · source-backed
Released September 18 by a lab not previously counted as frontier. Artificial Analysis scores it 44 on the Intelligence Index, matching Kimi K3 max, at $1.00 per million input and $2.70 per million output with a 95% cache discount, 99.8 output tokens per second, 1M context. Artificial Analysis Closed source, so unlike the recent Chinese open-weight releases there's nothing to run locally, which is the caveat on the "another Chinese lab reached the frontier" framing.
Each link below shares sources, entities, or timing with this story.
Moonshot exposes an Anthropic-compatible endpoint, so pointing Claude Code at K3 means setting the Anthropic base URL and supplying a Moonshot key. No new CLI, no config rewrite. Hosted at $3/$15 per Mtok, same tier as Claude Sonnet 4.6, and Artificial Analysis scores K3 at 57...
The v4.2 update published September 4 adds AA-Briefcase, a private eval of agentic knowledge work on expert-built projects, and GDP.pdf, a document-reasoning test over 4,592 pages of tables, charts and footnotes. GPQA Diamond was dropped as saturated. Fable 5.1 leads overall,...
Artificial Analysis has Grok 4.6 at 61, one point under Fable 5 Max's 62, leading GDPval-AA v2 at 1753 (vs 1741 and 1728) and AA-Briefcase at 1577 (vs 1574 and 1502), beating Sol on 6 of 9 shared benchmarks. Then Terminal-Bench v3.0: 26% versus 34.6% for Sol and 34.1% for Fabl...
OpenAI cut GPT-5.6 Luna roughly 80%, from $1 to $0.20 per million input and $6 to $1.20 output. Anthropic priced Opus 5 at $5/$25 per million, half of Fable 5. The trigger is DeepSeek, Zhipu's GLM-5.2 and Moonshot's Kimi K3 landing 60-90% below US flagship pricing, with DoorDa...
On Artificial Analysis's Intelligence Index, per-task cost lands only slightly below OpenAI's top model, roughly double GLM-5.2 and about 20× DeepSeek V4. The author attributes it to verbosity: more tokens spent per solved problem, with the penalty worse on office-work tasks t...
Everyone kept score wrong. When OpenAI shipped GPT-5.6 (the Sol flagship plus Terra and Luna) to GA on July 9, then xAI put out Grok 4.5, Meta dropped Muse Spark 1.1, and Cognition shipped SWE-1.7, the reflex was to ask who won the benchmark. Wrong question. On the Artificial...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.