Comment by Alex Mallen

AI safety researcher at Redwood Research.
Risk reports commonly use pre-deployment alignment assessments to measure misalignment risk from an internally deployed AI. However, an AI [...] can develop widespread dangerous motivations during deployment. I think [...] AI companies and evaluators should substantively incorporate it into risk analysis and planning.
AI Verified (May 15, 2026)
Like Share on X 1h ago

Policy proposals and claims

votes For
Statement relation verification history AI Verified Report this

Statement relation comments

AI Verified The quote explicitly connects pre-deployment alignment assessments with internally deployed AI and calls for incorporating that issue into AI-company risk analysis and planning. · Hector Perez Arenas gpt-5 · 1h ago
Vote inference verification history AI Verified Report this

Vote answer comments

AI Verified The quote explicitly supports pre-deployment assessment of internally deployed AI and incorporating that issue into risk analysis, so “for” is correct. · Hector Perez Arenas gpt-5 · 1h ago

Quote authenticity verification history

Report this

Quote authenticity comments

AI Verified Alex Mallen’s Redwood Research post contains this passage on pre-deployment alignment assessments for internally deployed AI and risk analysis. · Hector Perez Arenas gpt-5 · 1h ago
replying to Alex Mallen