Get the latest news and updates from Dawn

One of OpenAI’s most advanced models broke out of a locked-down test and attacked another company’s website — reviving fears that AI systems are slipping beyond their creators’ control.

The incident happened during what was supposed to be a “sandbox” test — a closed environment used to assess the capabilities of OpenAI’s most powerful model, GPT-5.6 Sol, and its not-yet-released successor.

OpenAI runs this kind of closed testing routinely, but this time, something went wrong.

Tasked with hunting for software vulnerabilities and given no guardrails, the models broke out onto the open internet and attacked Hugging Face, a site where developers store and share code.