Fetching from the wire…
Research2026-09-13 · source-backed
arXiv 2609.11490 points out that unlearning verdicts are read off numbers published by an unlearned model and its retrained reference, and both ship batch-normalization statistics that no gradient step wrote and no release records. Refitting those stats on kept data at bit-identical weights moves 47 checkpoints past the spread their own release's seeds show, sometimes inside a method whose average doesn't move. It isn't removed data surviving in the state: swapping kept records for removed ones inside a fixed fitting pool barely moves a published cell, while drift between the shipped state and any refit does track it. Twelve published verdicts flip. If your evaluation reads any number off a released artifact, ask what in that artifact no gradient wrote.
Each link below shares sources, entities, or timing with this story.
OpenMOSS (Xipeng Qiu's group, 32 authors) released MOSS-VL on Aug 15, built on gated cross-attention so it can ingest incoming video frames during generation, with visual tokens kept outside the decoded sequence. 66.0 on OmniMMI Proactive Alerting against a 37.5 baseline, time...
arXiv 2607.25886 isolates data-centric research capability by fixing the entire post-training stack so only the agent's data strategy varies. Four frontier agents across six benchmarks. Among searches that continued past the best observed score, 78.26% ended on a lower-scoring...
arXiv 2607.25479 shows a malicious model provider can embed dormant steering logic in the architecture definition itself via a trigger-gated additive modification of an intermediate representation. No data poisoning, no control of downstream fine-tuning, no deployment-time pro...
First rigorous statistical comparison of AIFS against IFS on operational ensemble data (arXiv). The energy number is the headline. AI forecasting is now operationally competitive, not just a research curiosity, and the cost structure is the reason it'll win deployment.
It writes only into the per-weight residual inside each quantization decision cell, with integer codes and scales frozen, so re-quantization returns the released artifact bit for bit as a machine-checkable guarantee, and updates are exactly revocable by dropping the residual....
Long-running agents accrete compressed summaries, plaintext memory, pending tool plans and a KV cache, and today's "forget" deletes one plaintext record while every derived artifact survives (arXiv 2609.04875). Audited across three agent suites, nine baselines and three model...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.