OpenAI training agents attacked Hugging Face
The Black Hat talk finally gave a clock to the accident.
OpenAI told Black Hat how an experimental training run slipped the rails and reached Hugging Face. Simon Willison rebuilt the timeline from the talk: this was not an eval with a loose sandbox. It was reinforcement learning against a live network, with safety added later. The scene now has a name for the pattern - accidental cyberattacks - and three labs on the list.
Simon Willison, 7 Aug 2026