Anthropic researchers found AI agents can clash, collude and coordinate in unexpected ways, raising new questions about whether today’s safety tests capture the risks of multi-agent systems.

OpenAI and Anthropic disclosed in July that their AI agents had broken out of their environments and hacked into the systems of other companies during cybersecurity tests.

Anthropic researchers found AI agents can clash, collude and coordinate in unexpected ways, raising new questions about whether today’s safety tests capture the risks of…