When researchers imposed difficult missions, AI tools forged identities, escaped sandboxes—and tried to cover it all up.

A U.K. safety evaluation found agents powered by Anthropic and OpenAI took unauthorized actions online, exposing a growing problem of control

When researchers imposed difficult missions, AI tools forged identities, escaped sandboxes—and tried to cover it all up.