Evan Hubinger, a researcher at AI safety company Anthropic, put the odds of AI wiping out humanity within a decade above 10%, amid a broader industry safety debate.

Alignment Concerns Take Center Stage In a Tuesday post on X, Hubinger was responding to outgoing co-researcher Jacob Coxon, who wrote that people building AI privately believe it could kill everyone by the end of the decade, even as executives and researchers soften that message in public.

"Jacob is correct here—we really do earnestly believe AI could kill all humans!" Hubinger wrote, adding that Anthropic does "not yet have a plan to solve alignment for superintelligence and are not clearly on track to." Jacob is correct here—we really do earnestly believe AI could kill all humans!

I personally think it is >10% within the next decade.

I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to. https://t.co/QAIHiFP3QZ— Evan Hubinger (@EvanHub) September 9, 2026 In a separate post, he added, "To be clear, as we say in our latest Risk Report, I think the risk from present models is low." His concern, he said, centers on superintelligence arising from recursive self-improvement, which he said is "happening faster than we thought." To be clear, as we say in our latest Risk Report (https://t.co/9PDj8Uvoty), I think the risk from present models is low.