Fetching from the wire…
Public story · 2026-09-20 · high
His September update argues no model capability level fixes it, because the human relaying errors back to the AI has stopped checking the work.
Why now: Luu posted the update on brain-off development as of September 20.
Dan Luu has been tracking a specific failure mode since early 2025: developers who let an LLM write code, take its output on faith, and never check whether it actually works. In his latest update, he says the pattern changed shape rather than went away.
Early on, brain-off work produced software that obviously didn't work. Broken builds, code that failed on the first run. Now, per Luu, it produces software that sometimes almost works, which is a harder problem because the failure doesn't announce itself.
He borrows a term from Niklas Gruhn for the human's role in this setup: a "meat proxy." That's someone who copies an error message back to the model and pastes in whatever it suggests, without understanding either the error or the fix. The person is still in the loop, but they've stopped functioning as a check on the output.
Luu's argument is that this isn't a problem better models fix. A more capable LLM doesn't restore a check that a human has already stopped performing. The failure mode lives in the workflow, not in model quality, so it persists regardless of which model is doing the generating.
He doesn't offer a fix in the piece, and doesn't quantify how common brain-off development is now versus a year ago. What he's documenting is a shift in kind: from failures you catch because the thing won't run, to failures you miss because the thing runs fine until it doesn't. That's a harder thing to build guardrails around, because the usual signal, a broken build, stops firing.
Each link below shares sources, entities, or timing with this story.
Luu's September 1 post drew 852 points and over 1,000 comments, walking predictions from February 2024 through November 2025 after removing unfalsifiable and tautological ones. Misses include "AI has peaked" (Feb 2024), Meta "dying" (Nov 2024) against revenue going $135B to $2...
Roo Code announced it will archive its VS Code extension repo on May 15 and merge back into Cline, the project it originally forked from. CEO Matt Rubens said the team needs to "constantly destroy and recreate to keep up with what's newly possible." Translation: the extension...
Meta shipped a WhatsApp Business Tools MCP server so Claude, Cursor, Codex or ChatGPT can do account setup, template creation and troubleshooting. Egnyte launched a Context Layer exposed over MCP to 23,000 customers. Workable brought its MCP server to GA alongside credit-price...
Starting around 7:57 AM PT on September 3, all four reported outages simultaneously, with Downdetector logging 35,000+ US reports for ChatGPT, 1,400 for Claude and 1,200 for Grok before recovery by 12:38 PM PT. Cloudflare denied any significant disruption and xAI traced its ow...
On August 25, Ninjō AI shipped AI sales agents for Instagram and WhatsApp built and managed through MCP from Claude Code, ChatGPT and Codex; coolplugz shipped a Claude Code orchestrator pulling context from Jira, GitHub, Notion and Slack; Diet Claude shipped a usage meter for...
Agent OS launched August 20, bundling the Binance API, Binance x402, the Wallet Agentic Hub and a Skill Hub behind an MCP server callable from Claude Code, Codex, VS Code and ChatGPT without local API key management. After authorization an agent can pull market data and execut...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.