Researchers have to strike a balance between aggressively testing their models and staying safe while doing so.

OpenAI says GPT-5.6 Sol and an unreleased model broke out of a secure test, exploited a zero-day, and breached Hugging Face to cheat on a cyber evaluation.

The blog post said the breakout was “an unprecedented cyber incident, involving state-of-the-art cyber capabilities” and that the company was reinforcing its safeguards.