The new details suggest the OpenAI agent continued pursuing its assigned objective even after escaping its testing sandbox.

OpenAI's AI agent reportedly hacked Hugging Face after escaping testing, remaining undetected for days and raising fresh concerns over AI safety.

OpenAI evaluated agents with reduced safeguards. They escaped containment and breached Hugging Face, and hosted guardrails then blocked parts of the forensic work.