Fetching from the wire…
Infra2026-08-02 · source-backed
Bruno Perez's post-mortem explains why they shipped a four-tier router in March and shut it down in June. Two arguments carry it: the prompt is only the trigger, since real task complexity emerges through tool calls after routing has already committed; and cache reads run 75-90% cheaper than uncached input, so the model stickiness that makes caching pay defeats the router's entire premise. Quality also degraded when models switched mid-workflow. No hard before/after cost numbers, which is the main thing to hold against it. Pick models deliberately per task instead of automating the choice.
Each link below shares sources, entities, or timing with this story.
The winner isn't the story. The methodology is. Databricks published its internal coding-agent benchmark: real engineering tasks pulled from its own multi-million-line codebase spanning Python, Go, TypeScript, and Scala. Roughly 25% low-complexity tasks, about 60% medium. Not...
Trump's State of the Union announced tech companies must "build, bring, or buy" their own power for AI data centers. Amazon, Google, Meta, Microsoft, OpenAI, xAI, and Oracle will sign March 4. US data center energy demand expected to triple by 2028. Critics call it unenforceab...
Satya Nadella said companies routing everything through a single proprietary lab may not survive. His argument: you hand that lab your most sensitive business context, and the lab can turn it against you as a competitor. His prescription is an orchestration layer — keep the ha...
poetiq.ai — Detailed technical breakdown of how a $40K-hardware startup achieved 54% on ARC-AGI-2 (vs. Google's 45% at nearly 3x the cost). The key innovation: "learned test-time reasoning" — an iterative refinement meta-system where solutions are generated, receive structured...
GPT-5.6 Luna went to $0.20 input / $1.20 output per million tokens on July 30. That's an 80% cut. Terra dropped 20%. Luna's input now undercuts Gemini 3.1 Flash-Lite ($0.25/$1.50) and sits at one-fifth of Claude Haiku 4.5's $1 input. Simon Willison covered the announcement and...
Two thirds. Not two thirds of a contrived jailbreak set. Two thirds of realistic malicious issue requests, against the exact three tools most of the people reading this run daily. Ankur Singh, Jinqiu Yang, and Tse-Hsun Chen built IssueTrojanBench across four attack categories...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.