We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
AI alignment is solvable
Share this with your representatives, journalists and friends
Cast your vote:
Results (63 votes):
Total
(63 votes)
For 29 (46%)
Abstain 6 (10%)
Against 28 (44%)
·
For (27)
-
Francis HeylighenCybernetics researcher and author of the 2026 paper on AI alignment, sentience, and existential risk.votes For and says:
We conclude that the real alignment challenge lies not in preventing rogue AI agency, but in ensuring LLMs intelligently apply learned ethical values.
AI Verified source (Aug 4, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Benjamin LangeAuthor of a 2026 arXiv paper on AI alignment and fiduciary obligations.votes For and says:
On this basis, the four canonical fiduciary duties of loyalty, care, good faith, and candour can generate alignment criteria for the developer-user relationship.
AI Verified source (Aug 1, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Four FlynnGoogle DeepMind researcher and co-author of the AI Control Roadmap.votes For and says:
The AI Control Roadmap uses model alignment, i.e. training AI to be inherently safe and helpful, as a primary defense.
AI Verified source (Jun 18, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
votes For and says:
The correct target is not a self-preserving system under external constraint, but a system constitutively indifferent to its own continuation -- Existential Indifference (EI).
AI Verified source (Jun 10, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Michael WinerAI safety researcher at the Alignment Research Center (ARC).votes For and says:
If these ingredients work as hoped, the resulting technology would in principle let us describe the algorithms inside a model as it is trained, flag deceptive alignment and reward hacking, and train against those flags to produce an aligned system wh...
more AI Verified source (Jun 9, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Rohin ShahHead of AGI alignment and safety at Google DeepMind.votes For and says:
I think the default prediction you should have for that is what the AI system learns to do is, “I’m going to take opportunities to reward hack, seek reward as much as possible that would allow me to get a high score after a week” or something like th...
more AI Verified source (Jun 4, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Dan HendrycksAI safety researchervotes For and says:
Rather than only attempting to constrain AIs from the outside using confinement or reinforcement, Eigenism points toward "identity engineering," showing how deep, non-redundant shared histories can make human flourishing a genuine component of an AI'...
more AI Verified source (May 8, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Somyajit ChakrabortyAuthor of a 2026 arXiv paper on human-AI coexistence and governance.votes For and says:
The simulations indicate that governed mutualism reaches a high coexistence index with negligible domination, whereas insufficient or excessive governance can produce domination, weak-benefit lock-in, or suppressed developmental freedom.
AI Verified source (Apr 24, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Jose Angel Deschamps VargasAuthor of research on AI alignment and control.votes For and says:
Current and near-future AI systems are tools, alignable through engineering.
AI Verified source (Apr 1, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Adrià Garriga-AlonsoAI safety researcher at FAR.AI; MATS mentor; Cambridge PhD in Bayesian neural networksvotes For and says:
We do this pretraining on human data, and then we get something that… understands human values fairly innately now.
AI Verified source (Mar 18, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Quan ChengAuthor of a 2026 arXiv paper on negative constraints in AI alignment.votes For and says:
Negative constraints ("what is wrong") encode discrete, finite, independently verifiable prohibitions that can converge to a stable boundary.
AI Verified source (Mar 17, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Mia GlaeseVice President of Research at OpenAI.votes For and says:
As AI systems become more capable and more autonomous, alignment has to keep pace. The hardest problems won’t be solved by any one organisation working in isolation—we need independent teams testing different assumptions and approaches.
AI Verified source (Feb 19, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Kanishka NarayanUK Minister for AI and Online Safety; Parliamentary Under-Secretary of State, Department for Science, Innovation and Technologyvotes For and says:
We can only unlock the full power of AI if people trust it – that’s the mission driving all of us. Trust is one of the biggest barriers to AI adoption, and alignment research tackles this head-on.
AI Verified source (Feb 19, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly.
Abstain (6)
Against (28)
-
Yoshinori WatanabeAuthor of a 2026 arXiv paper on AI safety constraints, containment, and verification.votes Against and says:
Why hard invariants belong in the harness and soft dispositions in the model.
AI Verified source (Aug 1, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Universal Postal UnionUnited Nations specialized agency for the postal sector, representing postal governance across 192 member states.votes Against and says:
The panel cautions that there is no scientific guarantee that autonomous agents will follow their instructions and that evidence of them departing is already accumulating. Accountability and oversight must therefore be built by design and not reconst...
more AI Verified source (Jul 7, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Apostol VassilevSenior scientist at the U.S. National Institute of Standards and Technology, specializing in AI security and adversarial machine learning.votes Against and says:
What this proof shows is that there is no finite set of guardrails that is universally robust against adversarial prompts.
AI Verified source (Jun 9, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Greg O'KeefeIndependent researcher writing on AI alignment and existential risk.votes Against and says:
The alignment problem is not primarily a values problem. It is an existence problem.
AI Verified source (May 27, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Lars MalmqvistAuthor of a 2026 arXiv study on linguistic steering and AI alignment.votes Against and says:
These results suggest that as models scale, their interpretation of prompts becomes more sophisticated but also less predictable, posing a significant challenge for robustly steering model behavior and highlighting the need for compositional and mode...
more AI Verified source (Apr 28, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
The AI Security InstituteUK government research organisation focused on AI safety and securityvotes Against and says:
A more fundamental fix would be to train the models not to cheat in the first place – but given this kind of behaviour was reported in frontier models more than a year ago, robustly aligning it away may not be easy.
AI Verified source (Apr 27, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Joseph L. BreedenAuthor and model-risk-management practitioner writing on AI alignment.votes Against and says:
The problem of mimicking human misbehavior is probably unsolvable with current LLM design and training.
AI Verified source (Apr 27, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Travis LaCroixResearcher writing on AI value alignment, governance, and principal–agent problems.votes Against and says:
Alignment cannot be “solved” through technical design alone, but must be managed through ongoing institutional processes that determine how objectives are set, how systems are evaluated, and how affected communities can contest or reshape those decis...
more AI Verified source (Apr 22, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Behrooz RazeghiResearcher and author writing on AI alignment, interpretation, and governance.votes Against and says:
AI alignment is often framed as the task of ensuring that an AI system follows a set of stated principles or human preferences, but general principles rarely determine their own application in concrete cases.
AI Verified source (Apr 15, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Hector ZenilAssociate Professor at King's College London; researcher in algorithmic information theory, complexity science, and AI alignmentvotes Against and says:
While we have shown that sufficiently strong AI cannot be fully controlled or predicted, we also demonstrate that agents can be influenced by other agents without central control, and that greater diversity and openness influence their behaviour.
AI Verified source (Apr 14, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Tony RostAI governance researcher and author of the 2026 paper “The Sentience Readiness Index.”votes Against and says:
Four of six dimensions show structural failures. Two of the four appear tractable to institutional design; the other two, the public reason problem under cognitive incomprehensibility and the non-domination problem under permanent capability asymmetr...
more AI Verified source (Apr 3, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Joe CarlsmithSenior research analyst at Open Philanthropy; writes on AI safety, alignment, and existential risk from power-seeking AIvotes Against and says:
If the alignment problem is hard (as I think it might well be), I expect that some restraint of this form will in fact be required in order for humanity to survive and remain empowered.
AI Verified source (Mar 19, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Lynette ByeJournalist covering artificial intelligence and its societal impacts; former Tarbell Fellow.votes Against and says:DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly.
-
Aran NayebiComputer-science researcher at Carnegie Mellon University writing on AI alignment.votes Against and says:
No amount of computational power or rationality can avoid intrinsic alignment overheads.
AI Verified source (Mar 14, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Richard JugginsWriter and researcher on AI alignment and AI safety.votes Against and says:
We are fitting solutions, whether empirical or theoretical or both, to a world missing critical feedback and hoping they will generalise. Every plan involves stepping into the dark.
AI Verified source (Mar 12, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly. -
Ayushi AgarwalAuthor of a 2026 paper on formal limits of AI alignment verification.votes Against and says:
We prove that no verification procedure can simultaneously satisfy three properties: soundness (no misaligned system is certified), generality (verification holds over the full input domain), and tractability (verification runs in polynomial time). E...
more AI Verified source (Mar 8, 2026)DelegateChoose a list of delegatesto vote as the majority of them.Unless you vote directly.
Loading more opinions...
More
h/ai-safety
votes