Tag
Evan Hubinger has been appointed as Alignment Science Lead at Anthropic, marking a significant role in advancing AI safety research.
Evan Hubinger, Anthropic's Alignment Science lead, states that he believes there is a >10% chance AI could kill all humans within the next decade, and that Anthropic lacks a plan to solve alignment for superintelligence.