Anthropic's Alignment Science lead says there is a ">10%" chance AI
Anthropic's Alignment Science lead says there is a ">10%" chance AI could kill all humans within the next decade and worries about recursive self-improvement

TL;DRAnthropic's top scientist estimates >10% chance AI kills humanity within decade.
Why it matters: Leading AI safety researcher openly acknowledges existential risk; signals alignment remains unsolved problem.
Anthropic's Alignment Science lead says there is a “>10%” chance AI could kill all humans within the next decade and worries about recursive self-improvement — Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
Read full articleSource: @evanhub · Opens in new tab