Fetching from the wire…
Public story · 2026-09-24 · high
A study of 6,774 merged agent PRs found 69.6% of the follow-up fixes came from the same agent that wrote the original.
Why now: Five different coding agents now generate enough merged-PR volume in popular repos to measure what happens after the merge, not just before it.
Agent-written pull requests draw follow-up fixes at 1.62 times the odds of human ones, a comparison of 6,774 merged agent PRs against 5,044 human PRs in the same repos found.
The gap matters for anyone reviewing agent code as if it were done once merged. Repos in the sample all carry at least 500 GitHub stars. These are production codebases accepting PRs from Codex, Copilot, Devin, Cursor and Claude Code, and fixing them more often than they fix human work.
69.6% of the follow-up fixes on an agent PR come from the same agent that wrote it. 76.4% of fix PRs are agent-authored across every commit. Agents return to patch their own mistakes more than humans return to patch a stranger's, according to the paper on arXiv.
A merged PR isn't the finish line for agent output the way it is for human output. Repos that treat an agent's merge as equally final to a human's are signing up for a second review pass more often than they'd expect. The odds say that pass should go back to the same agent that wrote the code, not a cold human reviewer.
Each link below shares sources, entities, or timing with this story.
arXiv 2609.17598 studies PRs from OpenAI Codex, Devin, GitHub Copilot, Cursor and Claude Code across 2,807 repositories (Dec 2024 to Jul 2025), combining AIDev with 58,792 cached GitHub API responses. Codex PRs were reverted 6.1% of the time against a human baseline of 11.5% (...
Researchers analyzed 61,837 GitHub Actions runs from 2,355 repos triggered by PRs from Claude, Devin, Cursor, Copilot, and Codex. Substantial differences in pass rates across bots. This is the first empirical data on how AI-generated code actually performs under real CI/CD con...
Stewardship moved to the Agentic AI Foundation under the Linux Foundation, with 30+ tools reading it natively: Claude Code, Copilot, Cursor, Codex, Gemini CLI, Windsurf, Devin, Aider, Amazon Q (BuildBetter). Claude Code reads AGENTS.md in addition to CLAUDE.md, which stays its...
The TypeScript project compiles a repository into an explorable, searchable graph you can ask questions against, with the explicit framing "graphs that teach beat graphs that impress." It targets Claude Code, Codex, Cursor, and Copilot. Directly relevant if your agent currentl...
wanshuiyin/HERO-Anti-OverDefense went from creation to 68 stars in a single day. HERO is Hashing, Edge cases, Rubrics, Overbuild, and the claim is that agent over-engineering isn't diffuse but falls into four recognizable shapes suppressible with a portable prompt contract acr...
A study of 3,171 GitHub repositories (2,660 multi-component setups, 511 skill collections) measured only byte-decidable defects in Claude Code, Cursor, Copilot and Codex configuration artifacts, validating every finding through independent re-derivation, an LLM adjudicator and...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.