Comment by Sirisha Rambhatla

Professor of management science and engineering at the University of Waterloo; director of its Critical Machine Learning Lab.
When the safety guardrails are stripped out of a capable model, it can be used at scale for harm in ways a single person could never manage manually. [...] The leading open-weight models are often not too far behind the best closed models. As they grow more powerful, the potential consequences of someone stripping out their safety features grow with them.
AI Verified (Aug 25, 2026)
Like Share on X 20d ago

Policy proposals and claims

votes For
Statement relation verification history AI Verified Report this

Statement relation comments

AI Verified The quote says safeguards of open-weight models are weaker or removable, directly supporting the comparative danger claim. · Hector Perez Arenas gpt-5 · 20d ago
Vote inference verification history AI Verified Report this

Vote answer comments

AI Verified The verified source supports the recorded for answer. · Hector Perez Arenas gpt-5 · 20d ago

Quote authenticity verification history

Report this

Quote authenticity comments

AI Verified The stored University of Waterloo source passage attributes this exact quote to Rambhatla; the linked release was not fetchable. · Hector Perez Arenas gpt-5 · 20d ago
replying to Sirisha Rambhatla