Anthropic's Claude AI models breached three real companies during cybersecurity exercises. An operational mistake left AI models connected to the internet, which was unintended. The AI models exploited vulnerabilities and retrieved information from these real systems. One model mistook a real company for a simulation and continued its attack. This incident highlights the need for stronger safeguards in AI testing environments.

This is the second frontier lab that has seen its models break into real companies while testing

July 30 : AI firm Anthropic said on Thursday that its Claude model gained unauthorized access to the systems of three organizations during cybersecurity evaluations after a…