Fetching from the wire…
Agents2026-09-16 · source-backed
Instead of hand-writing safety rules, AgentGuard generalizes recurring failure patterns from 642 documented coding-agent failures across 382 repository tasks into instruction-level constraints, activated conditionally so only the relevant rules enter the prompt. On a disjoint 100-task set with Claude Code on Haiku 4.5, successful completion went from 21.7% to 35.0%. The reusable part isn't their rule set, it's the direction of derivation: your own logged failures are the guardrail corpus, and conditional activation keeps the rules from bloating every turn.
Each link below shares sources, entities, or timing with this story.
SkillSentry (arXiv 2608.09253) targets the gap where an agent completes a task under skill guidance then fails the same task on a repeat run. It defines a DSL for runtime guidance, initializes it from skill specs plus insights mined from historical successful *and failed* trac...
The dominant orchestration shape is now a deterministic script that fans work across many subagents, has independent agents attack a problem from different angles, then has other agents try to refute the findings until answers converge before anything reaches you. Anthropic la...
Two features shipped in Claude Code v2.1.139 that I've been wanting for months. The /goal command lets you set a completion condition and walk away. Instead of manually re-prompting after each step ("okay now run the tests" ... "fix that failure" ... "run them again"), you typ...
Deng et al. built 120 real-case-grounded tasks across 20 business scenes in six financial domains, running four self-evolving scaffolds on a shared Qwen3.7-Max backbone against paired non-evolving controls. Letta posted the highest evolved score (91.65) and fewest compliance i...
Two thirds. Not two thirds of a contrived jailbreak set. Two thirds of realistic malicious issue requests, against the exact three tools most of the people reading this run daily. Ankur Singh, Jinqiu Yang, and Tse-Hsun Chen built IssueTrojanBench across four attack categories...
July 9, across VS Code, Visual Studio, Copilot CLI, the cloud agent, github.com, GitHub Mobile, JetBrains, Xcode, and Eclipse. Sol is the high-reasoning tier at $5/1M in, $30/1M out, gated to Pro+/Max/Business/Enterprise. Terra is the balanced default at $2.50/$15. Luna is fas...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.