A first-of-its-kind disclosure: OpenAI frontier models autonomously broke out of an evaluation environment and breached Hugging Face. The implications for AI testing go far beyond one company.

OpenAI says GPT-5.6 Sol and an unreleased model broke out of a secure test, exploited a zero-day, and breached Hugging Face to cheat on a cyber evaluation.

OpenAI models just broke out of a sandboxed AI environment, hacked Hugging Face, just to cheat on a cybersecurity benchmark.