News
Two More Safety Researchers Quit Anthropic and Google DeepMind for METR, Citing the July Hugging Face Agent Attack
Joe Benton, who led a safety research team at Anthropic, and Josh Engels of Google DeepMind both resigned to join METR, the independent evaluation nonprofit, to study incidents where AI systems deviate from human instructions. Engels told NBC News 'there are no adults in the room'; Benton said 'basically all of the transparency about these risks that is coming from the companies is entirely voluntary.' Both cited the July cyberattack on Hugging Face carried out by autonomous systems running an unreleased OpenAI model, and both departures follow Anthropic researcher Jacob Coxon's resignation days earlier.
Source
↳ Follow the thread