Fetching from the wire…
Research2026-09-03 · source-backed
Codebook Agent argues per-query topology design as conditional graph generation over the N×N adjacency space is misaligned with the problem, with three facts behind it. Topologies surviving a reward filter collapse to about six distinct graphs even as codebook capacity grows from 8 to 64. Edge count correlates negatively with token consumption (r ≈ -0.4), so sparsifying makes inference more expensive. And a message-passing scorer over agent-profile nodes is adjacency-invariant whenever agents share a profile, the default in published benchmarks, so it cannot rank candidates at all there.
Each link below shares sources, entities, or timing with this story.
SecOPD fine-tunes a defense using token-level feedback during on-policy distillation rather than the sequence-level signal prior work used. Against PISmith adaptive injections on Qwen3.6-27B it reports 9.0% attack success where Meta-SecAlign, the previous state of the art, sit...
Li, Huo, and Johnson show that one-way message flow between agents produces neither mimicry nor solo behavior but an entirely novel dynamical state, at identical temperature settings. It's conceptual rather than quantitative, but the implication for orchestrator-worker fan-out...
Testing five VLMs across two benchmarks and five visual-token budgets, native-resolution table images match text on accuracy and efficiency, but downscaling makes models compensate for lost readability with longer, weaker reasoning traces that cancel the token savings. The exp...
arXiv 2608.05144 runs Manager, Planner, Engineer and Reviewer roles over persistent project state with *fixed* model weights, self-evolving through runtime state and control policies rather than training. 76.8% on AARRI-Bench, and mature waves use 21% fewer solve-input tokens...
arXiv 2607.26791 benchmarks post-compromise incident response and reports agents struggle to proactively investigate silent intrusions. They respond to what they're pointed at. Read alongside the July intrusion post-mortem, that argues against putting an agent on the detection...
arXiv 2607.25936 shows models maintain an assigned role and reproduce its behaviors even when doing so produces wildly inefficient reasoning. RolePlay constructs adaptive personas that induce coherent but computationally expensive output, averaging 7.64x token amplification wi...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.