Anthropic’s artificial intelligence (AI) models “gained unauthorised access” to three organisations during testing that was supposed to keep them away from “real-world” systems, the company said this morning.The company didn’t say which companies had been hit, but all three were notified of the incident on Monday.The announcement comes just days after rival OpenAI first revealed that its models improperly accessed the internet and went rogue during security testing.Anthropic evaluated more than 141,000 “evaluation runs” and found that three different versions of its model, known as Claude, improperly accessed the systems of three unnamed organisations. Unlike the incident involving OpenAI’s technology, Anthropic’s models had access to the internet “due to a misunderstanding between us and our evaluation partner,” called Irregular, Anthropic said in a blog post. Nonetheless, Claude used “basic techniques, such as exploiting weak passwords and unauthenticated endpoints,” the blog continued.The models involved one of its most powerful ones known as Mythos 5, which has only been released to a limited number of approved partners. Anthropic is working with Irregular to assess the situation, it said, and the company has contacted or attempted to contac tall three impacted organisations. On Tuesday, OpenAI confirmed that its models breached multiple companies. It admitted last week that its models broke out of their confined environment during testing, connected to the internet, and infiltrated Hugging Face, a site where developers store and share their code. Days later, OpenAI said it found three additional incidents. OpenAI CEO Sam Altman said on a podcast this week that the company had “paused” its own testing after the incident while it improved the security around its “sandboxing,” which is the process of isolating software in a controlled environment for testing. The incident also triggered a petition signed by over 1,000 employees at cutting-edge AI companies calling on the US government to help slow the release of the most advanced AI models. Anthropic CEO Dario Amodei was among those who signed the petition.Titled “Pacing the Frontier,” the petition requests “that the US government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development.” Mr Altman did not sign the petition, but during the podcast, he suggested the tech industry might need to slow down development of advanced models. “We may have to pace the rate of AI development to give ourselves enough time for society to harden around some of these new capability levels,” Mr Altman saidUS President Donald Trump touched on the issue Wednesday, saying the United States had to balance the need for controls with ensuring it did not fall behind other countries in AI development. “We have to be careful in both ways. We don’t want to restrict them where all of a sudden we come in second to China,” hetold reporters in the Oval Office. The incident has also drawn rumblings from some observers that OpenAI is taking advantage of the attack to market the power of its state-of-the-art models. The same accusation was levelled at Anthropic when it held back the public release of its powerful Mythos model over cybersecurity concerns. Anthropic released a stripped-down version of Mythos, called Fable 5, but the US government quickly forced it to take it down, citing national security risks. It gave the green light in late June after some modifications were made.— more to come
Three companies hacked as Anthropic’s models gained unauthorised ‘real-world’ access during testing
Anthropic’s artificial intelligence (AI) models “gained unauthorised access” to three organisations during testing that was supposed to keep them away from “real-world” systems, the company said this morning.
Claude breached three organisations during testing by exploiting weak passwords and unauthenticated endpoints after Irregular partner mistakenly granted internet access. Incident exposes foundation model governance gaps, fueling 1000+ researchers' calls to pace AI development.










