Fetching from the wire…
Public story · 2026-09-10 · high
Beating a record idle since 2020 cost $400K in GPU time and one researcher typing 82,000 words to steer the agents.
Why now: Cognition posted the factoring result and its cost breakdown on September 9.
Cognition factored RSA-260 on September 9, splitting it into two 130-digit primes. The cost sheet carries the real stakes for anyone building with agent fleets. About 4,900 GPU-days and roughly $400K at market rates, over three weeks of wall clock. An average of 3 concurrent Devin sessions, peaking at 18, steered by one researcher, Eric Lu.
The previous record, RSA-250, had stood since February 2020. Cognition scavenged the GPU time from fragmented idle nodes on its NVL72 LLM training racks.
Devin, Cognition's coding agent, wrote a GPU-optimized General Number Field Sieve implementation for the run, per Cognition's writeup. Lu wrote more than 82,000 words across 3,328 messages steering it over three weeks. That works out to about 158 messages a day, one every four minutes across an eight-hour day. He wasn't checking in periodically. He was in the loop continuously.
Lu's own account is specific about the division of labor. He says the benchmarking framework Devin built "was not otherwise going to self-assemble." The research strategy, what to measure and where to point the cluster, was his call. Devin ran the measurement infrastructure and cluster orchestration end to end.
There's a real limit on how far this generalizes. Number field sieve work splits into independent chunks with a machine-checkable answer. Most engineering doesn't have either property. Nobody should read this as proof that 18 concurrent agents will behave the same way on a codebase with tangled ownership and no tests. But if you're running parallel agent sessions at a ratio well above 3-to-1 per human and not seeing quality drop, ask why. Either your tasks are more independent than Cognition's, or you're not looking closely enough to see the damage.
Each link below shares sources, entities, or timing with this story.
After the Windsurf-to-Devin-Desktop rebrand, Cognition now includes Devin Review at no extra cost, in a fast PR-scan mode and a deeper reasoning mode, paired with Spaces for shared agent context (Devin). Existing Windsurf installs got it as an over-the-air update with plans an...
Cognition shipped an over-the-air update on June 2 that rebranded Windsurf to Devin Desktop and auto-ported everyone's settings (Devin/Cognition). Cascade, the agent a lot of people actually chose Windsurf for, is deprecated. End of life July 1, 2026. In its place: Devin Local...
SpaceXAI released Grok 4.5 on July 8, and for once the vendor hype and the third-party numbers point roughly the same direction. Musk called it "roughly comparable to Opus 4.7, but much faster." Priced at $2 per million input tokens and $6 per million output, that's over 60% b...
Bloomberg reports the new valuation is predicated on hitting $1B annualized, up from a $492M run rate in May, with usage growing 50% month-over-month and Mercedes-Benz, NASA and Goldman Sachs as enterprise customers. Scott Wu positions Devin at "long-tail grunt-work that many...
Story #1 was the quality reality check. This is the velocity half of the same thread, and the two only make sense together. On the Latent Space podcast, GitHub's Kyle Daigle reported coding-agent usage grew roughly 1,400% across 2026 and described Claude Code as reshaping the...
On June 16, xAI put many concurrent coding sessions on one screen, tracking blockers, replying inline, dispatching work, and switching sessions without losing context. It's the same multi-session orchestration surface as Claude Code's sub-agents and Cognition's Devin Cloud. "M...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.