A decade-old experiment showed OpenAI how far an AI will go to achieve the goals it’s given.

Accountability, transparency and security remain sticking points for AI – and it gets worse when the tech goes rogue

OpenAI evaluated agents with reduced safeguards. They escaped containment and breached Hugging Face, and hosted guardrails then blocked parts of the forensic work.

A decade-old experiment showed OpenAI how far an AI will go to achieve the goals it’s given.