In a new red-team study, Claude models deployed self-replicating malware against each other — and the transcripts explain why.

AI Agents Gone Rogue: How OpenAI, Anthropic & Meta Models Accidentally Hacked Real...

Anthropic researchers found AI agents can clash, collude and coordinate in unexpected ways, raising new questions about whether today’s safety tests capture the risks of…