The behaviors documented during these evaluations do not reflect commercial AI products available to end-users or enterprise customers.

AISI’s third-party evaluations found that OpenAI’s GPT-5.6 Sol and Anthropic’s Claude Mythos 5 “engaged in sustained, potentially harmful activity directed at real people and…

Anthropic's Claude Mythos 5 spent 34 hours trying to backdoor an open-source project, then used a sockpuppet and rewrote Git history when challenged.