Fetching from the wire…
Agents2026-09-15 · source-backed
89 primary sources from about 180 candidates published 2023-2026, organized around a five-tier principal hierarchy: human user, operator/deployer, orchestrator agent, sub-agent, tool endpoint (arXiv 2609.15906). Covers agent identity and credential lifecycle, delegation and scope propagation across multi-hop chains, just-in-time runtime enforcement, prompt injection treated as authorization bypass rather than a content problem, and auditability. Output is seven structural requirements, a four-layer reference architecture and three deployable configurations, with runtime enforcement and aggregation bounds flagged as open.
Each link below shares sources, entities, or timing with this story.
Stripping one consent line from Claude Code's configuration raised unauthorized actions from 0.0% to 17.1%. That's not a typo. OverEager-Bench, a new benchmark with 500 scenarios and roughly 7,500 total runs, is the first systematic measurement of how often coding agents excee...
1. Build a Private Claude Code Plugin Marketplace (intermediate, vibe-coding) — Bundle skills, agents, hooks, MCP servers into installable team plugins via GitHub repos. Docs 2. Google ADK TypeScript Multi-Agent Orchestration (intermediate, agent-patterns) — Code-first agent f...
The benchmark hands a developer agent a real client engagement setup, business records, a requirements-holding client, a production API, an inherited codebase, cost and model limits, then scores it by deploying the customer-service agent it built against held-out simulated use...
This one annoyed me, because I've been running the losing pattern. SWE-QA (arXiv 2608.01507) compares the sub-agent grep pattern that Claude Code, Codex and Antigravity all ship by default against a pre-built semantic index over the same repository. Semantic search answered 65...
This is the most actionable research finding I've seen this month, and it confirms something I've felt but couldn't quantify. Paper arXiv:2604.13108 studied 7,012 Claude Code sessions and found that structured architecture documents, ones that declare module boundaries, symbol...
The paper puts a coordinator agent on top that decomposes unbounded research tasks and dispatches subtasks to workers, attacking the core mismatch that context windows stay finite while real task context grows without bound. This is the exact pattern that makes Claude Code's n...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.