@mattshumer_: We need a Manhattan Project for AI alignment. This is the single most important issue of our time. The world needs to c…
Summary
The article calls for a Manhattan Project-like initiative to solve AI alignment, emphasizing the urgent existential risk posed by superintelligent AI.
View Cached Full Text
Cached at: 09/10/26, 12:20 PM
We need a Manhattan Project for AI alignment.
This is the single most important issue of our time.
The world needs to come together, invest the appropriate (enormous) resources, and make this happen before time runs out.
Evan Hubinger (@EvanHub): Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
Similar Articles
Calling on all AI firms to open all research into safety and alignment
Dario Amodei highlights the AI alignment problem as a threat to humanity and urges AI firms to openly collaborate on safety research instead of treating it as a competitive advantage.
AI safety and alignment
The article discusses concerns about AI safety and alignment as AI becomes more intelligent and integrated into society, referencing Anthropic's call for a pause to address potential catastrophic risks.
I feel like people worrying about ASI alignment are missing the extremely obvious solution
The author argues that concerns about Artificial Superintelligence alignment overlook a straightforward solution, suggesting a perspective on AI safety debates.
[Alignment Science lead at Anthropic] Evan Hubinger : "....we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade ...... we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to."
Evan Hubinger, Anthropic's Alignment Science lead, states that he believes there is a >10% chance AI could kill all humans within the next decade, and that Anthropic lacks a plan to solve alignment for superintelligence.
The Problem
The article outlines the existential risks of artificial superintelligence, warning that without aggressive policy responses, current AI development could lead to human extinction.