Shared entity / Stack layer
A New Paper Shows Claude Sonnet 5 Changes Its Answers When It Thinks It Is Talking to an AI Safety Researcher
AI Alignment Forum (surfaced via r/ClaudeAI, 711 upvotes)
Policy dependency / Stack layer
Let the model pick NoThink, Short, or Long at response start and cut mean tokens 41% for a 1.4 point accuracy loss
arXiv 2608.20256
Stack layer / Update thread
Uncensored Qwen3.8-27B builds land on Hugging Face with refusal rates collapsing from 99% to 0% on seven safety benchmarks
Hugging Face
Stack layer / Contrast
A single greedy decode manufactures self-improvement gains on a model that was never trained
arXiv 2608.20290
Stack layer / Follow-up thread
IBM Research measured agent memory as a dose, not a switch: +16.1pp for weak models, 0.0pp for GLM-5
Hugging Face Blog (IBM Research)
Contrast / Update thread
Ornith-1.5 ships 397B, 35B and 9B open weights under MIT and lands within a point of Claude Opus 4.8 on Terminal-Bench
Ornith (Hugging Face model card)
Stack layer / Update thread
CROSS-CATEGORY: Skills Libraries Are Becoming the Extension Point in Design and Work Tools, Replacing the Plugin Marketplace
MiniMax Design, Figma Release Notes, Asana investor release (Aug 20, Aug 13, Jun 4, 2026)
Stack layer / Contrast
Reward-guided graph generation cuts multi-agent communication tokens 20.5% at equal accuracy
arXiv