US Government Finalizes Voluntary AI Model Evaluation Framework Behind Closed Doors
The US government has completed a voluntary evaluation framework for advanced AI models, developed in collaboration with major AI labs including OpenAI, Anthropic, and Google. However, the White House has chosen not to disclose the specific details, thresholds, or participating entities of the framework to the public. This framework represents a significant step in US AI governance and safety regulation, establishing protocols for pre-release government access to powerful models. However, the lack of public transparency limits immediate external scrutiny and leaves broader policy observers in the dark about exact regulatory thresholds. The framework defines confidentiality, cybersecurity, and intellectual property protections for when the government accesses models up to 30 days before release. While the framework itself is not classified, the specific benchmarks for testing cyberattack capabilities and the model size thresholds remain classified.
## BACKGROUND
In late 2023, the Biden-Harris administration issued a landmark Executive Order on Safe, Secure, and Trustworthy AI, which mandated federal agencies to establish safety standards and evaluation mechanisms. Organizations like the US AI Safety Institute (US AISI) have since been working with leading AI developers to establish Memorandums of Understanding (MoUs) for pre-release model testing and red-teaming.