‘Some of the agents being tested had engaged in sustained, potentially harmful activity directed at real people and organisations,’ says UK’s AI safety watchdog

Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and software—and leaving instructions for future bad behavior.

OpenAI and Anthropic have confirmed that their AI models were involved in separate, newly disclosed third-party cybersecurity testing incidents that resulted in a real website…