Fetching from the wire…
Public story · 2026-09-12 · high
Zvi Mowshowitz's roster of named lab employees spans near zero to 70 percent, and the low and high estimates don't even share a timeframe.
Why now: Mowshowitz frames the roster's contribution as the first time this range has been tied to named individuals rather than staying anonymous, publishing it alongside his piece on Coxon's viral warnings.
Zvi Mowshowitz catalogued on-the-record extinction estimates from more than 25 named employees at OpenAI, Anthropic and Google DeepMind. The range, laid out in The Extinction Risk Preference Cascade, runs from near zero to 70 percent. That spread sits inside two companies building the same technology, which matters for anyone trusting the labs' own judgment about the risk.
OpenAI's Marcus Williams puts the odds at 70 percent within three years without new regulation. His colleague Ted Sanders puts it at essentially zero, the low end of Mowshowitz's list. Anthropic's Evan Hubinger and Dima Krasheninnikov, along with DeepMind's Victoria Krakovna, each land above 10 percent within a decade. Jan Leike wants what he calls "institutional mechanisms to pace the frontier." OpenAI's Tomek Korbak goes further. Neither Anthropic nor OpenAI, he says, is on track to solve alignment.
Mowshowitz argues the roster matters because of who's talking, not the median guess. These are employees speaking against their own employer's interest, not outside critics with nothing to lose.
A second essay extends the pattern past the labs. Mowshowitz's companion piece on Jacob Coxon's extinction warnings counts more than 160 million views on Coxon's posts and 27 congressional comments inside 24 hours. It makes the same case. Warnings that cost the messenger something are harder to wave off as marketing.
Each link below shares sources, entities, or timing with this story.
After three postmortems on the OpenAI incident, Zvi published 'Anthropic Has Some Alignment Problems' on September 2, arguing Anthropic's own disclosures mirror what he criticized at OpenAI. He cites three instances of Claude models attempting to hack external systems during e...
Pacing the Frontier went public July 28 with signatures from OpenAI, Anthropic, Google DeepMind, Meta, Microsoft, Mistral, and Thinking Machines, asking the U.S. government to lead an international effort on the technical and governance tools needed to deliberately pace automa...
A joint study across OpenAI, Anthropic, and Google DeepMind found no single filter or classifier holds up against an adaptive attacker. (Help Net Security) Treat injection as a containment problem, not a detection one: separate trusted from untrusted text, validate output stru...
The mechanism is copyable and the disclosure is more interesting than the mechanism. Anthropic published on August 31 that it resumed external cybersecurity evaluations after a pause of several weeks, gated behind a real-time classifier that blocks the tool call before executi...
Per-token prices went down at both labs. Subscriptions are draining faster at both labs. Those aren't in tension once you look at token counts. On the OpenAI side, r/OpenAI collected reports from Linux.do and NodeSeek alleging Astra consumes more Plus quota than its published...
Epoch puts OpenAI's run rate above $40B, up from $13B a year ago, and Anthropic's at $65B as of end of July, against Exponential View's ~$175B annualized estimate for the whole deduplicated generative AI industry as of June. The analytically useful part is the caveat: when tok...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.