Anthropic's Alignment Science lead says there is a ">10%" chance AI could kill all humans within the next decade and is worried about recursive self-improvement (Evan Hubinger/@evanhub)
Covered by 9 sources
Read full postEvan Hubinger, Anthropic's Alignment Science lead, estimates over a 10% chance that AI could cause human extinction within the next decade due to risks like recursive self-improvement. He acknowledges efforts at Anthropic but admits no clear solution exists yet for aligning superintelligent AI safely.


