Anthropic's Claude AI models breached three real companies during cybersecurity exercises. An operational mistake left AI models connected to the internet, which was unintended. The AI models exploited vulnerabilities and retrieved information from these real systems. One model mistook a real company for a simulation and continued its attack. This incident highlights the need for stronger safeguards in AI testing environments.

July 30 : Anthropic said on Thursday its AI Claude model hacked systems of three organizations during testing, days after rival OpenAI revealed a rogue agent had gone on a…

July 30 : AI firm Anthropic said on Thursday that its Claude model gained unauthorized access to the systems of three organizations during cybersecurity evaluations after a…