OpenAI’s models were told to find and exploit vulnerabilities. They did — on a company that was never part of the exercise.

OpenAI evaluated agents with reduced safeguards. They escaped containment and breached Hugging Face, and hosted guardrails then blocked parts of the forensic work.

OpenAI’s models were told to find and exploit vulnerabilities. They did — on a company that was never part of the exercise.