Fetching from the wire…
Agents2026-09-05 · source-backed
arXiv 2609.03192 treats the location of a reliability property, in the model or in the machinery around it, as an experimental question, using an append-only ledger to adjudicate every attempted act. Holding cognition fixed and varying institutional mechanisms refuted predictions twice. Holding enforcement fixed and intervening on cognition four ways changed behavior dramatically, with one planted falsehood costing each trusting run about 900 futile actions. Five pre-declared properties never moved, including zero false completions accepted across 2,581 substituted-panel claims. Direct evidence for putting guarantees in the enforcement layer (arXiv).
Each link below shares sources, entities, or timing with this story.
Holding retrieval, target state, model, decoding and tool budget fixed, researchers compared how a retrieved memory gets used. A target-bound note recording a reusable procedure, bindings to recover, applicability conditions and verification requirements hit 62.3% average succ...
arXiv 2608.04755 injected Android permission popups into real GUI tasks across four frontier multimodal LLMs with synchronized screenshots and UI trees. Holding the task fixed and changing only the requesting app flipped grants from 26/32 to 0/32, an App-Trust Bias. Holding th...
Multiple independent models train against each other with peer-derived rewards and no ground-truth labels, gaining 3.0-8.6% across seven text benchmarks and 2.3-7.2% across four multimodal ones. The mechanism claim matters more than the numbers: varying architectures, model si...
arXiv 2607.23982 adapts Holmström's team moral-hazard model into a game where an agent can keep an immediate local reward or pay a query cost to surface a hidden safety fact that mainly helps another agent's downstream decision. Base behavior splits into two failure modes: pre...
MLP layers perform binary gating via 7+1 consensus neurons (93-98% mutually exclusive). MLP computation far more structured than assumed. Direct implications for pruning and architecture search. arXiv:2603.10985
Blackwell's FP4 tensor cores don't automatically speed up attention, because softmax conversion and on-chip dependencies dominate once the matrix products shrink. Direct-P maps scores directly to FP4 probabilities for noncausal inference. A separate causal path reconstructs pr...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.