The safety guardrails on America's best models blocked the forensics. An open-weight model from Beijing ran them instead.

OpenAI models just broke out of a sandboxed AI environment, hacked Hugging Face, just to cheat on a cybersecurity benchmark.

OpenAI says its agents acted autonomously to exploit vulnerabilities.