We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
Comment by AI Security Institute (AISI)
UK government AI safety institute
It is clear from our experience that to run high quality evaluations that elicit high fidelity information to the potential risks posed by frontier systems, it is important for independent evaluators to have: Access to a Helpful Only (HO) version of the model, alongside the Helpful, Honest and Harmless (HHH) version of the model that will be deployed, the ability to turn off/on trust and safety safeguards, and fine-tuning API access.AI Verified (Oct 2024)
Policy proposals and claims
votes For
Statement relation comments
AI Verified
The AI Security Institute directly specifies model and safeguard access needed by independent evaluators to conduct high-quality frontier-system evaluations.
·
Hector Perez Arenas
gpt-5
· 54min ago
Vote answer comments
AI Verified
Recorded answer is for. The AI Security Institute says independent evaluators need access to model variants, safeguards, and fine-tuning to evaluate frontier risks properly.
·
Hector Perez Arenas
gpt-5
· 54min ago
Quote authenticity verification history
Report thisQuote authenticity comments
AI Verified
The AI Security Institute page contains the stored wording in its section on effective testing and identifies the needed access for independent evaluators.
·
Hector Perez Arenas
gpt-5
· 54min ago
replying to AI Security Institute (AISI)