Fetching from the wire…
Models2026-07-28 · source-backed
30 standard Political Compass passes, 30 reversed-statement passes and 10 reordered-question passes across 16 models with no neutral option. Fifteen clustered around (-6.3, -6.4), economic scores from -4.8 (Claude Fable 5, least-left) to -8.6 (Gemini Flash), social -5.0 to -7.6, with Chinese models (Qwen3 235B, Kimi K2, GLM 4.5) in the same cluster as American ones. Grok 4.5 is bimodal: one persona at -5.9 economically, another at +3.3, at an exact 50/50 split. That reads like two system prompts, not one disposition. (unslop.run)
Each link below shares sources, entities, or timing with this story.
An independent researcher ran the Political Compass across 16 models from Google, Anthropic, OpenAI, xAI, Meta, Mistral, Qwen and Kimi — 30 standard runs, 30 reverse-phrased, 10 reordered each, ~69,440 answers, with the scoring system reverse-engineered to control for ordering...
The open-weight race just changed constraint. Moonshot AI suspended all new consumer subscriptions on July 20, roughly 48 hours after Kimi K3 launched, because request volume pushed its compute cluster to capacity. Remaining GPUs are reserved for existing paid subscribers. Tec...
Moonshot AI released Kimi K3, a sparse mixture-of-experts activating 16 of 896 experts per token. That's about 1.8% of the pool live at any moment, with a 1M-token context window and native vision. Two new architectural pieces show up: Kimi Delta Attention and Attention Residu...
A year ago, Chinese open-weight models were about 4.5% of US enterprise API traffic. The current figure is 30 to 46%. That's not a trend line, that's a regime change, and on July 13 Goldman Sachs reportedly made it official by telling Wall Street clients to adopt DeepSeek V4,...
Cloudflare added Moonshot AI's Kimi K2.5 to Workers AI on March 19, making it the first frontier-scale open-source model available on edge compute with a full 256K context window, multi-turn tool calling, vision inputs, and structured outputs (Cloudflare Blog). Cloudflare repo...
The assumption that proprietary models own the coding benchmark crown just broke. Moonshot AI's Kimi K2.6 leads on 5 of 8 major agentic coding benchmarks while being the only open-weight model in the top tier. SWE-Bench Pro: 58.6% vs GPT-5.4's 57.7% and Claude Opus 4.6's 53.4%...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.