Back to News

Anthropic alignment lead: >10% chance AI kills all humans in a decade

45 points · 100 comments#anthropic#ai-safety#alignment#existential-risk

Anthropic's Alignment Science lead Evan Hubinger stated there is a greater than 10% chance AI could kill all humans within the next decade, and that Anthropic lacks a plan to solve alignment for superintelligence. He expressed concern about recursive self-improvement and said the company is not clearly on track to address the issue.

Coverage timeline

  1. Techmeme

    Evan Hubinger / @evanhub : Anthropic's Alignment Science lead says there is a “>10%” chance AI could kill all humans within the next decade and worries about recursive self-improvement — Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.

  2. Hacker Newsljf