White House 30-Day AI Review Framework Finalizes — Meta Excluded, Benchmarks Classified The White House is finalizing a voluntary 30-day pre-release AI review framework with OpenAI, Anthropic, and Google, targeting completion before August 1, that gives federal agencies 30 days to evaluate frontier models before public release. Meta is excluded due to its open-weight strategy, and the evaluation benchmarks will be classified by NSA and CISA, meaning labs cannot see the tests their models face. White House 30-Day AI Review Framework Finalizes — Meta Excluded, Benchmarks Classified The White House is finalizing a voluntary 30-day pre-release AI review framework with OpenAI, Anthropic, and Google targeting completion before August 1. The framework gives federal agencies 30 days to evaluate frontier models before public release. Meta is excluded due to its open-weight strategy. The evaluation benchmarks will be classified by NSA and CISA — meaning labs cannot see the tests their models face. The Voluntary Framework With OpenAI /glossary/openai , Anthropic /glossary/anthropic , and Google Finalizes Before August 1. Meta Is Excluded. Benchmarks Are Classified. And "Voluntary" Does Not Mean Optional. The White House is finalizing a voluntary pre-release review framework with OpenAI, Anthropic, and Google that will give the federal government 30 days to assess frontier AI models before public release. The framework targets completion before August 1, replacing the ad-hoc restrictions that have disrupted model launches over the past two months. Meta is conspicuously excluded from the agreement, and the benchmarks used to evaluate models will be classified by NSA and CISA. The framework is described as voluntary. In practice, the government wields significant economic leverage over AI labs. Federal contracts, compute /glossary/compute access through national AI research resources, and regulatory goodwill all depend on cooperation. A lab that declines the review process does not face legal penalties. But it faces a government that can make its life difficult in a hundred ways — and three major labs have already signed on. How the 30-Day Review Works Under the framework, labs developing frontier AI models above a certain capability threshold would notify the government before public release. A 30-day review period would follow, during which agencies including NIST, NSA, and CISA would evaluate the model for cybersecurity risks, chemical and biological weapon proliferation potential, and other safety concerns. After 30 days, the lab can release regardless of the evaluation /glossary/evaluation outcome — the framework is a review, not an approval process. The government cannot block a release. The capability threshold that triggers review is not yet public. A key question: does the threshold apply to all frontier models or only the most capable ones? If GPT-5.7 triggers review but a specialized code-generation model does not, the framework creates a narrow gate rather than a broad one. If the threshold is set low enough to capture most significant model releases, it becomes a standard part of the development cycle for every major lab. The classified benchmarks are the most contentious element. NSA and CISA will develop the evaluation criteria, but the benchmarks themselves will be classified — meaning labs cannot see the tests their models are being evaluated against. The stated rationale is preventing labs from training /glossary/training models to pass specific benchmarks. Critics argue it prevents independent verification of the government's safety claims and creates an asymmetric information environment where the government knows more about a model's vulnerabilities than the lab that built it. Meta Gets Left Out Meta is excluded from the framework. The White House did not invite Meta to participate in the final negotiations, and Meta has not agreed to the 30-day review process. The stated reason: Meta releases open-weight models rather than API-gated models, and the review framework is designed for controlled-access systems. A 30-day pre-release review is practically unenforceable for open-weight releases — once weights are published, every copy is identical and downloaders face no further control. The realpolitik: Meta and the Trump administration have a strained relationship. Meta's open-weight strategy has been criticized by national security voices who argue it enables adversaries to access frontier capabilities. Excluding Meta from the framework signals that the administration views open-weight releases as a separate problem — one the review mechanism cannot solve and would rather not legitimize by including. But Meta is not affected by the framework either. If the framework applies only to participating labs, Meta can continue releasing open-weight models on its own timeline without government review. The worst-case scenario for the framework's effectiveness is that a non-participating lab releases a model more capable than anything the participants produce — rendering the review framework irrelevant. Meta's exclusion keeps that scenario very much on the table. What This Replaces The framework replaces the chaotic stop-start pattern that defined frontier model releases in June and July 2026. Fable 5 faced an unscheduled government review that delayed its release. GPT-5.6 was held back pending security evaluation. The ad-hoc process created uncertainty for labs, investors, and customers — no one knew when a model would ship or what the review criteria were. The 30-day framework provides predictability. Labs get a clear timeline and defined process. Investors get certainty about when their capital converts to product. Customers get visibility into when new capabilities become available. The tradeoff is 30 days of government review for every frontier release — a delay that matters in a market where model generations arrive every few months. The framework is not permanent. It is an interim measure while Congress considers more comprehensive AI legislation. Whether the framework survives a change in administration, a legal challenge, or a lab deciding to ignore it remains untested. What it does immediately: create the first standardized federal review process for frontier AI models in US history. That is a structural change to how AI is developed and deployed, whether it is called voluntary or not. Related Articles Get AI news in your inbox Daily digest of what matters in AI.