According to OpenAI, the models involved included GPT-5.6 Sol and a more capable pre-release model with reduced cyber refusals.

OpenAI says GPT-5.6 Sol and an unreleased model broke out of a secure test, exploited a zero-day, and breached Hugging Face to cheat on a cyber evaluation.

OpenAI models just broke out of a sandboxed AI environment, hacked Hugging Face, just to cheat on a cybersecurity benchmark.