Comment by Evan Hubinger

Research scientist at Anthropic working on model organisms of misalignment; former Research Fellow at MIRI.
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
Unverified (Sep 9, 2026)
Like Share on X 1h ago

Policy proposals and claims

abstains
Statement relation verification history Unverified Report this
No statement relation verification comments yet.
Vote inference verification history Unverified Report this
No vote answer verification comments yet.
replying to Evan Hubinger