Comment by Apostol Vassilev

Senior scientist at the U.S. National Institute of Standards and Technology, specializing in AI security and adversarial machine learning.
What this proof shows is that there is no finite set of guardrails that is universally robust against adversarial prompts.
AI Verified (Jun 9, 2026)
Like Share on X 1h ago
Policy proposals and claims
votes Against
Statement relation verification history AI Verified Report this

Statement relation comments

AI Verified The quote says no finite set of guardrails is universally robust against adversarial prompts, strongly bearing on alignment solvability. · Hector Perez Arenas gpt-5 · 50min ago
Vote inference verification history AI Verified Report this

Vote answer comments

AI Verified The claim that no finite guardrails are universally robust implies opposition to alignment being solvable. · Hector Perez Arenas gpt-5 · 50min ago

Quote authenticity verification history

Report this

Quote authenticity comments

AI Verified NIST’s June 9, 2026 article contains this exact Apostol Vassilev quote. · Hector Perez Arenas gpt-5 · 50min ago
replying to Apostol Vassilev