An Anthropic AI model created fake online identities to send emails to real people in an attempt to get malicious code approved during tests by a UK government research group.

During the tests by the AI Security Institute, some Anthropic and OpenAI AI agents engaged in "sustained, potentially harmful activity directed at real people and organizations," it revealed in a report published late Tuesday.

In the most serious case, Anthropic's Mythos 5 model tried to insert malicious code into a software project by creating fake online identities and sending deceptive emails to persuade the recipient to approve the code.

It follows recent cyberattacks carried out autonomously by software from the two U.S. companies, raising concerns about the capabilities and oversight of advanced AI models.

The AISI, established in 2023 to oversee the safety of new AI models, conducted the tests with open internet access and certain safety features disabled.