Fetching from the wire…
Vibe Coding2026-09-09 · source-backed
Retrieval returns flat lists of isolated snippets, so agents pick a semantically similar sibling. RepoNav is a post-retrieval interface presenting compact structural cues and candidate targets, guiding on-demand file-structure browsing, and it improves function-level localization on LocBench across models. Ablations attribute the gain to structured organization rather than to simply exposing more file structure. arXiv 2609.07925
Each link below shares sources, entities, or timing with this story.
arXiv 2607.29658 attacks the fact that repair agents treat every issue independently and throw away procedural knowledge. STAIR converts historical trajectories into multi-level trees spanning fine-grained diagnostic actions up through high-level strategies, then tailors plan...
LatentMD separates content correctness from boundary correctness in CommonMark fence handling across 9 LLMs and about 37,600 generations. Ablations attribute failures primarily to same-family symmetric-delimiter collisions, not nesting depth, and the problem generalizes to Pyt...
Sergey Rodionov's paper tests four Codex-based agent variants to isolate what actually drives performance. Verification (simplification plus exact observation reproduction) ranked highest in every setting, but at substantially higher cost. The textual baseline beat the executa...
The FSE '26 paper argues SWE-bench, SWT-bench, and AgentBench capture narrow synthetic slices, and proposes contamination-aware, trajectory-aware, in-the-wild evaluation using agents' commit signatures to study real vs human contributions over time. (arXiv) Pair this with the...
Attention-observed selectors like H2O and SnapKV collapse to 0.00-0.33 needle retrieval on a NoPE MLA model, because a long-lived cache must be compressed before the queries that will read it exist. On Kimi Linear, VestigeKV evicts by a query-independent signal already in the...
July 17, Product Hunt's #1 product was Unabyss for Claude: shared memory across all apps and LLMs, 598 votes. July 18, #1 was ZooData: "the data layer for AI agents," 606 votes. Neither is an application. Both are substrate. (Product Hunt) One launch is noise. Two consecutive...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.