Fetching from the wire…
Models2026-09-03 · source-backed
Released September 2, Quasar 438B scores 43 on the Artificial Analysis Intelligence Index, 75.0 on AA-LCR long-context reasoning and 69.3 on Terminal-Bench v2.1, with a 15.3-second response time for 500 tokens. That's above Mistral Medium 3.5 at 30 and NVIDIA Nemotron 3 Ultra at 38, and well behind Claude Opus 5 at 63. English and Spanish only, served exclusively through the CompactifAI API, no weights or license disclosed.
Each link below shares sources, entities, or timing with this story.
Moonshot AI dropped Kimi K2.6 today and the numbers are hard to ignore. One trillion parameters total, 32 billion active per token across 384 experts, 256K context window, and native multimodal input. It scores 58.6 on SWE-Bench Pro versus GPT-5.4's 57.7 and Claude Opus 4.6's...
xAI launched Grok 4.5 and Grok Build on July 8, trained partly on Cursor developer-session data. The numbers are loud: 83.3% on Terminal-Bench 2.1, 64.7% on SWE-Bench Pro, priced at $2/$6 per million tokens. On a single coding task that works out to roughly $2.49 versus $11.80...
Shopify launched an open-source AI Toolkit on April 9 that does something no major commerce platform has done before: it gives external AI coding agents full operational control over real stores. Not read access. Not sandboxed previews. Full control. Build apps, update product...
Cursor stopped being an IDE wrapper and became a model company. Cursor shipped Composer 2, a proprietary coding model trained via reinforcement learning on long-horizon coding tasks. On CursorBench — their own benchmark, caveats acknowledged — it scores 61.3, beating Claude Op...
I don't care that Grok 4.5 ranks #4. I care that it resolves a SWE-Bench Pro task with an average of 15,954 output tokens where Opus 4.8 spends 67,020. That's a 4.2x efficiency gap, and it lands straight in my monthly bill. SpaceXAI launched Grok 4.5 on July 8, a roughly 1.5T-...
NVIDIA launched Nemotron 3 Super — a 120B total / 12B active parameter hybrid Mamba-Transformer MoE, open, designed specifically for multi-agent workloads, and delivering 5x higher throughput than Nemotron 2 at the same active parameter count (NVIDIA Newsroom). It ships with a...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.