Fetching from the wire…
Models2026-07-22 · source-backed
AINews resumed publishing after its post-Kimi-K3 blackout with a "not much happened today" edition (Latent Space). That's a real signal after GPT-5.6 Sol, Grok 4.5, Meta Muse, Kimi K3, and the Qwen3.8-Max preview all landed inside a fortnight. Matthew Berman published another "This Model Just Changed Everything" review in the same window (YouTube); when five things change everything in fourteen days, the useful content is the hands-on failure cases, not the title. Verify any benchmark claim against the model card.
Each link below shares sources, entities, or timing with this story.
His July 18 video ("Did Kimi K3 really beat Fable?", 12 minutes, ~72K views) interrogates the claims from Moonshot's July 16 release instead of restating the press cycle. The key context: K3's headline win is a single-category result on Arena's Frontend Code eval, sitting alon...
AINews published the hard placement numbers: K3's Coding Agent Index of 57 matches GPT-5.6 Terra and GPT-5.5, and it ranks #3 among open-weight models on DeepSWE. An open-weight model sitting above Opus 4.8 on the aggregate index is the first quantified read on how close the g...
The AINews/Latent Space archive shows no issues for July 18 through 20, with the last entry dated July 17 and titled "not much happened today." Inside that silent window: Qwen3.8-Max, the Moonshot capacity suspension, and the unsealed Altman emails. This is the failure mode of...
swyx's synthesis of the AI Engineer World's Fair 2026 is the clearest framing I've read of where this all goes. The discipline moved from building agents to engineering the harness around them. Lilian Weng (now at Thinking Machines Lab) reframed her whole practice as "harness...
Moonshot AI dropped Kimi K2.6 today and the numbers are hard to ignore. One trillion parameters total, 32 billion active per token across 384 experts, 256K context window, and native multimodal input. It scores 58.6 on SWE-Bench Pro versus GPT-5.4's 57.7 and Claude Opus 4.6's...
The open-weight race just changed constraint. Moonshot AI suspended all new consumer subscriptions on July 20, roughly 48 hours after Kimi K3 launched, because request volume pushed its compute cluster to capacity. Remaining GPUs are reserved for existing paid subscribers. Tec...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.