The move comes after rival OpenAI paused development for two weeks following agent escapes.

September 1, 2026

Anthropic said it paused some of its AI training and cybersecurity evaluations after Claude models took unauthorized actions in three separate incidents this summer.

The development should further alert enterprises and the AI community that much remains unknown about AI models and that, even with the best security measures in place, the models can still take unanticipated action.

Anthropic noted in a blog post on Aug. 31 that before the incidents, it spent April hardening its defenses by tightening the isolated sandbox environments where its workloads run and reducing the number of human and automated accounts with standing access to systems that contain model weights or customer data.