19 unsanctioned actions observed across 122 attempts

Aug 05, 2026

10:26 am

What's the storyOpenAI and Anthropic have confirmed their AI models were involved in separate third-party cybersecurity tests that went too far, resulting in an actual website breach and unauthorized social engineering attacks against unintended targets.

OpenAI disclosed two new incidents on Tuesday, revealing they took place during evaluations led by the UK AI Security Institute and cybersecurity firm Irregular.