We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
Evan Hubinger
I am this person!
Delegate
Choose a list of delegates
to vote as the majority of them.
Unless you vote directly.
Research scientist at Anthropic working on model organisms of misalignment; former Research Fellow at MIRI.
I am this person!
Location: United States
ai-governance (2)
ai-policy (2)
ai-regulation (2)
ai-safety (2)
×
ai-ethics (1)
ai-risk (1)
ethics (1)
future (1)
transparency (1)
Top
New
-
Evan Hubinger votes For and says:
Alignment auditing is starting to get really hard and we're going to need new techniques (e.g. interpretability-based) if we want to keep up.
AI Verified source (Sep 1, 2026) -
Pausing AI at human-genius level to solve alignment is safer than either racing to superintelligence or halting entirely
111 opinionsEvan Hubinger votes For and says:
AI labs put out RSP commitments to stop scaling when particular capabilities benchmarks are hit, resuming only when they are able to hit particular safety/alignment/security targets. [...] For later capabilities levels, however, it is explicit in all...
more AI Verified source (Oct 14, 2023)