Comment by Sophie Kim

Author of The Counterfactual, publishing policy analysis on AI safety and biosecurity.
For closed-weight models, [misuse] risk is at least partially manageable. Input-output classifiers and safety fine-tuning can intercept most non-sophisticated actors before they get actionable guidance. [...] Open-weight models have no such layer. Once weights are downloaded, they can be run locally with no oversight and no logging, safety fine-tuning can be stripped with minimal technical effort, and the model can be fine-tuned on domain-specific data to become dramatically more capable in exactly the areas we’d most want to restrict. There is no API to monitor, no company to refuse the request, and no recourse once the weights are out.
Unverified (Apr 11, 2026)
Like Share on X 1h ago

Policy proposals and claims

votes For
Statement relation verification history Unverified Report this
No statement relation verification comments yet.
Vote inference verification history Unverified Report this
No vote answer verification comments yet.
replying to Sophie Kim