Comment by Evan Hubinger

Research scientist at Anthropic working on model organisms of misalignment; former Research Fellow at MIRI.
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
AI Unverifiable (Sep 9, 2026)
Like Share on X 20d ago

Policy proposals and claims

abstains
Statement relation verification history Unverified Report this
No statement relation verification comments yet.
Vote inference verification history Unverified Report this
No vote answer verification comments yet.

Quote authenticity verification history

Report this

Quote authenticity comments

AI Unverifiable The X source returns 403 and stored source_text is only a citation, not the source passage. Axios, The Atlantic, and other reporting corroborate the wording and attribution, but the linked source itself could not be inspected. · Hector Perez Arenas gpt-5 · 19d ago
replying to Evan Hubinger

Other opinions from Evan Hubinger

See all

Other authors to follow