White House's Closed-Door Approach to AI Safety Testing
Here's what actually happened. The recent executive order on AI safety and the accompanying framework being developed by various agencies has been deliberately scoped to exclude open models—those with publicly available weights—from mandatory capability testing requirements. This isn't buried in dense legalese either; it's a fairly explicit carve-out that anyone following the policy discussions would notice.
Why does this matter? Open models like Meta's Llama series, Mistral's releases, and the countless fine-tuned variants built on top of them represent a significant portion of real-world AI deployment. Startups, researchers, and even enterprise teams are building production systems on these foundations. Yet when it comes to the government's approach to governing advanced AI capabilities, these models are essentially operating in a policy blind spot.
The rationale, as I understand it from the policy discussions, is that you can't meaningfully restrict what's already out there. Once weights are public, traditional oversight mechanisms become largely theoretical. But this feels like a missed opportunity rather than a practical necessity.
Consider the alternative: what if the framework required documentation and reporting standards for models above certain capability thresholds, regardless of whether they're open or closed? We're already seeing companies voluntarily share safety data—Anthropic's recent CASA framework submission comes to mind. A government nudge toward transparency, even for open models, could accelerate industry-wide best practices without stifling innovation.
The bigger picture here is about maintaining competitive advantage. The U.S. government's approach seems to be: regulate the controlled environments (major tech companies with closed models) while letting the open ecosystem innovate freely. It's a calculated bet that open development will produce breakthroughs faster than centralized control, and that the benefits outweigh the risks.
Whether that's the right call remains to be seen. But excluding open models from safety frameworks entirely feels like governing by omission rather than design.