Fetching from the wire…
Research2026-09-05 · source-backed
arXiv 2609.03376 targets organizations outsourcing vector indexes to untrusted clouds, where each query touches corpus-scale state, so a naive secure implementation costs minutes and about 90 GB of communication per query at million-document scale, and recent optimized systems still need 10-22 seconds. Spruce learns compact binary codes preserving candidates for full-precision reranking, turning corpus-wide embedding scoring into Hamming-distance computation under two-server MPC, with a corpus-calibrated fixed-radius protocol avoiding multi-round candidate selection. Across four corpora of 383K to 5.42M documents it preserves original search quality (arXiv).
Each link below shares sources, entities, or timing with this story.
arXiv 2608.00765 compresses retrieved docs into query-conditioned visual representations, sidestepping the trade-off where hard compression is query-aware but weak and soft compression is strong but needs costly offline encoding. Beats both baselines across varying retrieval d...
Agent-Orchestrated Adaptive RAG (arXiv:2606.05658) finds agentic enhancements are not universally beneficial. Dynamic query decomposition gained +0.17 MRR on a structured DevOps benchmark but degraded ranking precision on a multi-hop benchmark, and the self-reflective loop onl...
RAGAS-style evaluation checks correctness against a frozen snapshot, which means routine document updates and corrections can silently break production without moving a dashboard. This ASE 2026 paper defines 11 mutation operators perturbing at both the pre-chunk index level an...
The paper names a failure mode called query dominance, where a high-capacity encoder lets the query dominate the latent state and renders retrieved evidence functionally irrelevant (arXiv 2608.16776). GRIP imposes capacity asymmetry: full-dimensional decoder access to the quer...
"Less Context, Better Agents" shows that long-horizon, tool-using agents drowning in verbose tool responses complete tasks reliably with just a recency window and a compact running summary (arXiv:2606.10209). This is a useful corrective to the reflex of bolting RAG onto everyt...
The study extracted 130 clean atomic state transitions from 707 real issues in SWE-bench Lite and Verified. Plain RAG scored 0.57-0.59 answer accuracy; an LLM reranker didn't help and added latency, about 18 seconds against 2.1. A (subject, relation, object) supersession memor...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.