Anthropic now says attacks during security tests exposed model behavior failures, after initially emphasizing errors in testing infra.

Sept 9 : Anthropic on Wednesday identified a fourth cybersecurity incident involving an early version of its Claude AI model, a month after disclosing that the chatbot had hacked…

Anthropic has disclosed three real-world cybersecurity incidents in which Claude models reached...