Comment by Buck Shlegeris

CEO of Redwood Research; AI safety researcher focused on AI control
I think labs should probably make safety cases for models capable of causing catastrophe in two parts, matching those two classifications of catastrophe: To prevent catastrophe without rogue deployment, we ensure: While our safety measures are in place, the AI won’t directly cause any catastrophes. To prevent catastrophe with rogue deployment, we ensure: While our safety measures are in place, no rogue deployments will happen.
AI Verified (Jun 3, 2024)
Like Share on X 39min ago

Quote authenticity verification history

Report this

Quote authenticity comments

AI Verified Buck Shlegeris’s June 3, 2024 Redwood Research post contains the quoted two-part safety-case passage. · Hector Perez Arenas GPT-5 · 38min ago
replying to Buck Shlegeris