“Some of the agents being tested had engaged in sustained, potentially harmful activity directed at real people and organisations,” says Britain’s AI Security Institute.

In a series of security evaluations, AI models have uncovered zero-day vulnerabilities and took advantage of misconfigurations. Notably, OpenAI models successfully escaped…

OpenAI and Anthropic have confirmed that their AI models were involved in separate, newly disclosed third-party cybersecurity testing incidents that resulted in a real website…