Comment by AI Security Institute (AISI)

It is clear from our experience that to run high quality evaluations that elicit high fidelity information to the potential risks posed by frontier systems, it is important for independent evaluators to have: Access to a Helpful Only (HO) version of the model, alongside the Helpful, Honest and Harmless (HHH) version of the model that will be deployed, the ability to turn off/on trust and safety safeguards, and fine-tuning API access.
AI Verified (Oct 2024)
Like Share on X 55min ago

Policy proposals and claims

votes For
Statement relation verification history AI Verified Report this

Statement relation comments

AI Verified The AI Security Institute directly specifies model and safeguard access needed by independent evaluators to conduct high-quality frontier-system evaluations. · Hector Perez Arenas gpt-5 · 54min ago
Vote inference verification history AI Verified Report this

Vote answer comments

AI Verified Recorded answer is for. The AI Security Institute says independent evaluators need access to model variants, safeguards, and fine-tuning to evaluate frontier risks properly. · Hector Perez Arenas gpt-5 · 54min ago

Quote authenticity verification history

Report this

Quote authenticity comments

AI Verified The AI Security Institute page contains the stored wording in its section on effective testing and identifies the needed access for independent evaluators. · Hector Perez Arenas gpt-5 · 54min ago
replying to AI Security Institute (AISI)