Fetching from the wire…
Research2026-08-13 · source-backed
arXiv 2608.12253 shows the standard practice of training a policy against a single LLM simulating the user fails because the simulator is itself mode-collapsed, so the policy learns to exploit its dominant mode. Verbalized Sampling recovers up to 9% held-out success; Population Co-Training reaches 14%, with a human study confirming a similar gain on real users. They released SCOPE as open source. Same shape as the Anthropic conformity finding: one distribution sampled repeatedly is not diversity.
Each link below shares sources, entities, or timing with this story.
LLM uses OpenAI / Shared entities / Same source domain / Earlier coverage / Tension
Linked by a graph relationship (LLM uses OpenAI); both cover LLM, Scope; reported by the same outlet (arxiv.org).
LLM uses OpenAI / Shared entities / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (LLM uses OpenAI); both cover Anthropic, Same; overlapping topics (against, anthropic).
Anthropic released Claude / Shared entities / Same source domain / Earlier coverage / Tension
Linked by a graph relationship (Anthropic released Claude); both cover Anthropic, LLM; reported by the same outlet (arxiv.org).
LLM uses OpenAI / Shared entities / Earlier coverage
Linked by a graph relationship (LLM uses OpenAI); both cover Anthropic, Same, Scope; earlier Anthropic coverage from 2026-08-05.
Anthropic released Claude / Shared entity: LLM / Same source domain / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (Anthropic released Claude); both cover LLM; reported by the same outlet (arxiv.org).
Anthropic released Claude Code / Shared entity: Anthropic / Same source domain / Shared topic / Earlier coverage
Linked by a graph relationship (Anthropic released Claude Code); both cover Anthropic; reported by the same outlet (arxiv.org).
Simon Willison released LLM / Shared entities / Same source domain / Earlier coverage
Linked by a graph relationship (Simon Willison released LLM); both cover Anthropic, Same; reported by the same outlet (arxiv.org).
LLM uses OpenAI / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (LLM uses OpenAI); both cover Anthropic, Same; overlapping topics (anthropic, policy).