We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
Comment by Sophie Kim
Author of The Counterfactual, publishing policy analysis on AI safety and biosecurity.
For closed-weight models, [misuse] risk is at least partially manageable. Input-output classifiers and safety fine-tuning can intercept most non-sophisticated actors before they get actionable guidance. [...] Open-weight models have no such layer. Once weights are downloaded, they can be run locally with no oversight and no logging, safety fine-tuning can be stripped with minimal technical effort, and the model can be fine-tuned on domain-specific data to become dramatically more capable in exactly the areas we’d most want to restrict. There is no API to monitor, no company to refuse the request, and no recourse once the weights are out.Unverified (Apr 11, 2026)
Policy proposals and claims
votes For
No statement relation verification comments yet.
No vote answer verification comments yet.
replying to Sophie Kim