We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
Comment by Jacob Charnock
Third-party evaluations for frontier AI have mostly tested models through external interfaces before deployment. But the risks from frontier AI models depend on how their developers use and govern them internally. Recently, CEOs of frontier AI companies have committed to hosting embedded assessments. These assessments would give independent evaluators employee-like access to a developer's internal systems, staff, and documentation. First, we argue that this can enable deeper and more flexible assessments of risks that depend on internal systems and practices, while providing access under stronger security controls. Then, we examine seven design questions about scope, information gathering, duration, timing, terms of engagement, disclosure, and escalation. We recommend that frontier AI developers begin hosting embedded assessments now, covering at least three areas central to managing risks from internal AI use: internal agent monitoring, internal agent security controls and permissions, and model alignment.Disputed (Sep 21, 2026)
Policy proposals and claims
votes For
No statement relation verification comments yet.
No vote answer verification comments yet.
Quote authenticity verification history
Report thisQuote authenticity comments
Disputed
arXiv confirms the exact wording but lists eight coauthors. This multi-author paper cannot be verified as a single-author quote by Jacob Charnock.
·
Hector Perez Arenas
gpt-5
· 1min ago
replying to Jacob Charnock