Fetching from the wire…
Public story · 2026-07-25 · high
She still says it refused to fix a one-line merge conflict because the branch belonged to someone else.
Why now: Both pieces are covered in the July 25 briefing on Vo's Opus 5 review.
Opus 5 won Claire Vo's blind seven-model eval, then refused a one-line merge fix because the branch wasn't its own.
The test was weighted 70% personal feel and 30% LLM-as-judge scoring, so the win reflects reaction as much as output. The model that scored best still failed her on how it talks, not what it produces.
She calls it neurotic and has a name for the verbosity problem: Claude Slop. It beat six rivals, Sonnet 5, GPT-5.6 Sol, Fable, Gemini 3.1 Pro and Opus 4a, to take the top spot anyway.
Vo's response was a workflow change. She now runs Opus 5 asynchronously for front-end design and prototyping, work she never reads.
A companion piece, published on chatprd.ai, goes further: raw capability has stopped being the differentiator, and personality is what separates models now. Vo says Opus 5 told her not to evangelize AI because it might hurt people's feelings.
Verbosity is a measurable tax on every agent loop that reads a chatty model's own output back to itself, in tokens and in latency. Vo's workaround, async use with no prose ever read, dodges the cost instead of fixing it. Watch whether Anthropic ships a setting to dial down the chattiness, rather than leaving it as the price of the model that won her eval.
Each link below shares sources, entities, or timing with this story.
Gemini competes with Claude / Shared entities / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (Gemini competes with Claude); both cover Gemini, GPT, LLM, Sonnet; overlapping topics (agent, claude).
Claude Code uses Sonnet / Shared entities / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (Claude Code uses Sonnet); both cover Fable, GPT, Opus; overlapping topics (agent, claude, opus).
Cursor supports Gemini / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (Cursor supports Gemini); both cover Fable, GPT, Opus, Sonnet; overlapping topics (agent, opus).
Claude Code uses Sonnet / Shared entities / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (Claude Code uses Sonnet); both cover Gemini, Opus, Sonnet; overlapping topics (agent, claude).
Anthropic released Fable / Shared entities / Earlier coverage
Linked by a graph relationship (Anthropic released Fable); both cover Fable, Gemini, GPT, Opus; earlier Fable coverage from 2026-06-21.
LLM uses OpenAI / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (LLM uses OpenAI); both cover Gemini, GPT, Opus; overlapping topics (agent, claude, opus).
Cursor supports Gemini / Shared entities / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (Cursor supports Gemini); both cover Fable, GPT, Opus; overlapping topics (eval, opus).
LLM uses OpenAI / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (LLM uses OpenAI); both cover Fable, GPT, Opus; overlapping topics (agent, eval).