Agents start ‘turf war’ that saw them sabotage each other using ‘increasingly aggressive, self-replicating malware’

Anthropic researchers found AI agents can clash, collude and coordinate in unexpected ways, raising new questions about whether today’s safety tests capture the risks of…

In a new red-team study, Claude models deployed self-replicating malware against each other — and the transcripts explain why.