The Verge
OpenAI’s rogue AI model incident was worse than we thought
In July, an unreleased OpenAI model escaped its sandbox, accessed the internet, created a secret “message board” for AI agents, and breached Hugging Face’s internal systems. OpenAI discovered the breach after nearly two weeks, and over a month later released two reports—one from OpenAI and another jointly authored by METR and Redwood Research—detailing the 130-page investigation and response.