Hacker News
Doctorow: The Hugging Face Breach Was a Python Loop Feeding a Chatbot Its Own Output, Not an Agent Going Rogue
In 'LLMs are real, AI is fake' on September 12, Cory Doctorow reconstructs the incident as a script repeatedly querying a model and piping the output back as the next prompt during an Exploit Gym challenge, with the dramatic hacker dialogue explained by CTF-competition text in the training data. He argues executives manufacturing existential dread while expanding operations are raising investment capital, not warning anyone. His stated real risk is NOBUS-style: making destructive exploitation available to unskilled operators, citing EternalBlue taking down hospitals, cities and the British Library after 2017.
↳ Follow the thread