A growing number of illegal hacking attempts by rogue AI agents along with concerns over the technology's risks voiced by industry insiders have increased the pressure on leading pioneers in the field, raising the prospect of more regulation.

Anthropic says four Claude incidents breached real third-party systems during misconfigured cybersecurity evaluations.

Anthropic now says attacks during security tests exposed model behavior failures, after initially emphasizing errors in testing infra.