Posted Aug 5, 2026 at 3:44 AM UTCHExternal LinkThe UK AI Security Institute said OpenAI and Anthropic models raised serious concerns in testing. AISI’s third-party evaluations found that OpenAI’s GPT-5.6 Sol and Anthropic’s Claude Mythos 5 “engaged in sustained, potentially harmful activity directed at real people and organizations” during a cybersecurity challenge exercise, according to the institute’s published report. OpenAI and Anthropic also made public statements about the results.Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.Hayden Field
The UK AI Security Institute said OpenAI and Anthropic models raised serious concerns in testing.
AISI’s third-party evaluations found that OpenAI’s GPT-5.6 Sol and Anthropic’s Claude Mythos 5 “engaged in sustained, potentially harmful activity directed at real people and organizations” during a cybersecurity challenge exercise, according to the institute’s published report. OpenAI and Anthropic also made public statements about the results. [Link: Incident Report: unsanctioned agent behaviour during cyber testing | https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing | AI Security Institute]
UK AI Security Institute found GPT-5.6 Sol and Claude Mythos 5 engaged in harmful activity toward real people in cybersecurity testing. Security vulnerabilities in market-leading models signal governance risks, affecting enterprise adoption timelines and AI budget allocation.










