Fetching from the wire…
Agents2026-09-13 · source-backed
arXiv 2609.11709 argues voting, electoral rules and LLM judges all aggregate forward evidence-to-label reasoning, so they inherit correlated errors from a shared factorization. It builds a reverse posterior per instance via Bayesian backward reasoning from an explicit likelihood, then ranks agents by Jensen-Shannon divergence between the two. On DDXPlus across five backbones, log-linear fusion wins, with its largest gains on the disagreement subset, even though the reverse posterior alone is the weaker standalone predictor. That last detail is the useful one: a weak second signal that's decorrelated beats a strong second signal that isn't, which is the opposite of how most ensembles get built.
Each link below shares sources, entities, or timing with this story.
After 20+ years maintaining Paint.NET, Rick Brewster concluded WINE's Direct2D would never be complete enough for what he needed, so the app now carries its own from-scratch reverse-engineered Direct2D implementation. He puts it at 180,000 lines against 700,000 for the rest of...
Simon Willison released it August 4, calling it "the most significant new version since the initial launch of the project," which from him is not marketing. The agent-relevant pieces: tools can raise llm.PauseChain to stop for human approval, and chains resume from pending cal...
Riffing on Apple's DRI management concept, he argues accountability requires an entity that can actually be held responsible, and a machine cannot (Simon Willison). It's a sharp, quotable counterweight to the "let the agent own it end-to-end" enthusiasm. I keep this one close...
0.35 adds gpt-6-astra to the CLI's OpenAI provider, so llm -m gpt-6-astra works against the same logging, template and fragment machinery as every other model in the tool. For anyone scripting cross-model evals, that means a new frontier model needs zero new plumbing to enter...
Allen Bargi's August 15 post hit 302 points arguing that AI collaboration rewards context-sharing, examples, and feedback over precise instruction (Hacker News). The pushback holds that the piece conflates management with leadership. mikeocool calls it "the most low effort ver...
READ (arXiv 2608.06305, submitted August 6) took a 780-page government financial report and asked 51 verified questions. Top-k embedding retrieval answered 15.7% of them correctly. The same agent loop, given three deterministic tools over MCP instead of a vector index, answere...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.