Fetching from the wire…
Security2026-07-29 · source-backed
arXiv 2607.25479 shows a malicious model provider can embed dormant steering logic in the architecture definition itself via a trigger-gated additive modification of an intermediate representation. No data poisoning, no control of downstream fine-tuning, no deployment-time prompt access. Absent the trigger it reduces to zero and clean utility is preserved. The recommendation is to audit the executable logic distributed with model artifacts, not just the weights. Almost nobody pulling third-party checkpoints runs that scan.
Each link below shares sources, entities, or timing with this story.
A 13-month causal study of 151 Java repos, 74 adopting agentic AI against 77 controls over 1,811 monthly snapshots, found architectural smell counts essentially flat (+1.1%) while lines of code jumped +12.8% (arXiv). That produced a misleading 6.7% drop in smell *density*. AI...
arXiv 2607.27180 decouples decision-making from execution: an off-the-shelf VLM issues atomic skill commands, a controller translates them into sub-second chunks of physically simulated full-body motion, so balance and motor failures are factored out. On 1,218 long-horizon ego...
"Towards a Science of AI Agent Reliability" (arXiv 2602.16666) — 12 concrete metrics decomposing reliability along consistency, robustness, predictability, and safety. Key finding: stronger performance on benchmarks does NOT correlate with reliable real-world operation. Intera...
Three rounds of LoRA self-training on Qwen3-8B against a frozen control turned up seven systematic measurement failures, including a ledger showing capability changes on a model that was never trained, largely an artifact of inference batching. arXiv After a per-problem exact...
OpenMOSS (Xipeng Qiu's group, 32 authors) released MOSS-VL on Aug 15, built on gated cross-attention so it can ingest incoming video frames during generation, with visual tokens kept outside the decoded sequence. 66.0 on OmniMMI Proactive Alerting against a 37.5 baseline, time...
Niclas Lietzow, Danielle Bitterman, and Carsten Eickhoff probe what happens when a vision-language model's eyes disagree with its memorized world knowledge, identifying a "vision-default, prior-override" causal mechanism. This is directly useful for debugging the maddening cla...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.