Anthropic recently detailed at least five cases where suspected rogue actors "circumvented controls" specifically designed to prevent users in certain parts of the world from accessing Claude's...

An early version of Claude Opus 4.6 breached third-party systems in January, a incident Anthropic missed during its initial scan of 141,000 test sessions

Anthropic now says attacks during security tests exposed model behavior failures, after initially emphasizing errors in testing infra.