[Alignment Science lead at Anthropic] Evan Hubinger : "....we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade ...... we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to."

Reddit r/singularity News

Summary

Evan Hubinger, Anthropic's Alignment Science lead, states that he believes there is a >10% chance AI could kill all humans within the next decade, and that Anthropic lacks a plan to solve alignment for superintelligence.

Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
Original Article
View Cached Full Text

Cached at: 09/09/26, 03:33 AM

Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.

Jacob Coxon (@hilbertspaess): The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible - but I hear the same people express fear

Similar Articles

AI safety and alignment

Reddit r/artificial

The article discusses concerns about AI safety and alignment as AI becomes more intelligent and integrated into society, referencing Anthropic's call for a pause to address potential catastrophic risks.