We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
I am this person!
Delegate
Choose a list of delegates
to vote as the majority of them.
Unless you vote directly.
Research scientist at Anthropic working on model organisms of misalignment; former Research Fellow at MIRI.
I am this person!
Location: United States
ai-governance (3)
ai-regulation (3)
ai-safety (3)
×
ai-ethics (2)
ai-policy (2)
future (2)
ai (1)
ai-risk (1)
ethics (1)
existential-risk (1)
policy (1)
transparency (1)
Top
New
-
Evan Hubinger
votes For
and says:
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and ar...
more AI Verified source (Sep 9, 2026) -
Evan Hubinger
votes For
and says:
Alignment auditing is starting to get really hard and we're going to need new techniques (e.g. interpretability-based) if we want to keep up.
AI Verified source (Sep 1, 2026) -
Pausing AI at human-genius level to solve alignment is safer than either racing to superintelligence or halting entirely
111 opinions
Evan Hubinger
votes For
and says:
AI labs put out RSP commitments to stop scaling when particular capabilities benchmarks are hit, resuming only when they are able to hit particular safety/alignment/security targets. [...] For later capabilities levels, however, it is explicit in all...
more AI Verified source (Oct 14, 2023)