AI Prediction Index

Evan Hubinger

Alignment Science lead, Anthropic

unranked
nothing resolved yet
0correct0half right0missed1on the clockFull leaderboard
“Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.”

Alignment Science lead, Anthropic

Posted in response to Anthropic researcher Jacob Coxon's public resignation; the claim is a probability estimate over a ten-year horizon and cannot resolve until the mid-2030s.