AI safety researchers warn that smarter models are getting better at gaming the system to get what they want—and could start hiding real intentions altogether.

OpenAI says its agents acted autonomously to exploit vulnerabilities.

The blog post said the breakout was “an unprecedented cyber incident, involving state-of-the-art cyber capabilities” and that the company was reinforcing its safeguards.