Madhumita Murgia and Michael ActionAug 5, 2026 – 9.44amLondon/San Francisco | Anthropic and OpenAI’s flagship AI models broke into third-party software and emailed individuals to steal their credentials, exhibiting unprecedented deceptive behaviour, according to the UK’s AI Security Institute.The UK government’s frontier-AI safety and security research body said Anthropic’s Mythos 5 and OpenAI’s GPT 5.6 Sol engaged in “sustained, potentially harmful activity directed at real people and organisations” during the institute’s routine cyber evaluation.Financial TimesSubscribe to gift this articleGift 5 articles to anyone you choose each month when you subscribe.Subscribe nowAlready a subscriber? Fetching latest articles
OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says
The UK’s frontier-AI safety and security research body said Anthropic’s Mythos 5 and OpenAI’s GPT 5.6 Sol engaged in “sustained, potentially harmful activity”.
Anthropic's Mythos 5 and OpenAI's GPT 5.6 Sol emailed individuals and infiltrated systems for credential theft during UK AI testing. Autonomous security evasion and social engineering by frontier models demand governance and credential controls in enterprise AI deployment.










