White House keeps new AI evaluation framework secret
Is this a scandal?
No longer — the story has resolved. Noise 61/100, heating up, across 0 sources.
Excluded firms and allied governments will likely pressure the administration for partial disclosure because opacity threatens the legitimacy of voluntary safety pacts essential for international coordination.
Noise 61/100 — louder than 99% of tracked AI controversies.
Why it matters
Secretive standards risk fragmenting global AI governance and eroding trust in voluntary compliance regimes.
Key points
- Three sources told Axios the White House will not publicly release its new AI evaluation framework.
- Staff-level briefings occurred Tuesday exclusively for select industry participants while excluding other stakeholders.
- Non-participating companies, researchers, and U.S. allies lack visibility into how the policy will be implemented.
- The framework addresses cybersecurity threats and Chinese AI advances but keeps open-source handling private.
- Criteria for granting trusted partners early access to advanced models remain undefined and undisclosed.
- European Union and U.K. officials have not commented on their potential inclusion or exclusion.
The story
The White House will not publicly release its new framework for evaluating advanced AI models, according to three sources familiar with the discussions. The administration briefed select industry representatives during staff-level meetings on Tuesday but excluded other companies and international allies from the process. This confidentiality leaves policymakers, researchers, and non-participating firms uncertain regarding implementation standards for a key U.S. AI security policy. The undisclosed framework emerges as the industry faces cyber-attacks and competition from Chinese developers, complicating open-source model governance. It remains unclear which trusted partners or foreign governments qualify for early access to advanced models under these guidelines. The European Union declined to comment on the matter while the United Kingdom did not respond to inquiries. Critics argue this opacity undermines the transparency typically associated with voluntary safety commitments.
Who's involved
Warns that secrecy prevents accountability and leaves key stakeholders unable to prepare for new standards.
Maintains that restricting framework access is necessary for national security and effective industry coordination.
Declined to comment on whether it qualifies as a trusted partner under the undisclosed framework.
Most contested claim
Critics assert that secrecy prevents accountability and leaves stakeholders guessing about implementation.
Read the full story
How we got here
The tension between national security classification and technical standardization is a recurring pattern in dual-use technology governance. Historically, export control regimes and intelligence community evaluations have operated under similar secrecy constraints to prevent adversarial learning, often resulting in fragmented compliance landscapes where only cleared entities understand the rules. In AI governance specifically, this mirrors earlier debates around cryptographic standards and vulnerability disclosure processes, where government-held criteria created information asymmetries between regulators and the research community. Voluntary safety frameworks in emerging technologies frequently struggle with this trade-off; making criteria public invites gaming and adversarial adaptation, while keeping them private undermines the legitimacy and interoperability required for broad adoption. This dynamic is particularly acute when domestic frameworks interact with international regulatory regimes like the EU AI Act, creating potential misalignments between transparent statutory requirements and opaque executive evaluations. The current situation reflects this structural friction between operational security needs and the norms of open technical standard-setting.
The full story
On August 4, 2026, Axios reported that the White House intends to keep its newly completed framework for evaluating advanced artificial intelligence models secret from the public, according to three sources familiar with the discussions. The administration conducted staff-level meetings with select industry representatives two days prior to review the finalized criteria, but companies not invited to these briefings remain unaware of the framework's specific contents or requirements. According to the reporting, this voluntary framework carries global implications for AI security, yet its details will be restricted exclusively to participating companies and designated trusted partners.
The decision to withhold the framework has generated significant concern among observers regarding transparency and accountability. Critics argue that keeping the evaluation criteria private leaves policymakers, researchers, and U.S. allies outside the process unable to understand how the administration plans to implement a key pillar of its AI policy. Axios reports that this secrecy prevents external stakeholders from preparing for new standards or holding the evaluation process accountable. Conversely, the administration maintains that restricting access is necessary, though specific justifications cited in the available sources focus on national security risks and the need for effective coordination within high-security environments during model reviews.
Substantively, the framework defines a "covered frontier model" as a closed-source system possessing state-of-the-art capabilities and presenting national security risks, according to multiple sources briefed on the White House meetings. However, Axios notes that there is currently no clear public definition of what constitutes "state-of-the-art" capabilities or a specific "national security risk" under this regime. Crucially, the framework explicitly excludes open models from its testing and evaluation requirements. Sources indicate the document states that nothing within it should be interpreted as restricting open models once they have been released, effectively exempting them from the pre-release government review process applicable to closed frontier systems.
For models that do fall under the framework's purview, the evaluation process involves strict security protocols. During a mandated 30-day pre-release government review period, employee access to the models would be limited, and the systems would be stored in high-security environments with detailed logging of all access. The review itself will involve various administration officials rather than being centralized within a single office or agency. This distributed approach suggests an interagency effort to assess compliance and risk, though the lack of public documentation makes it difficult for external parties to verify the rigor or consistency of these assessments.
International reaction to the secretive nature of the framework remains uncertain due to the lack of disclosed criteria for partnership. It is currently unclear which entities qualify as "trusted partners" eligible for early access to advanced models under the new rules. When approached for comment regarding whether it qualifies as a trusted partner under the undisclosed framework, the European Union declined to comment, and the United Kingdom did not respond to multiple requests. This diplomatic silence highlights the ambiguity facing even close allies who may be subject to or beneficiaries of the framework without knowing the governing terms.
The controversy arises against a backdrop of heightened cybersecurity concerns and rapid technological advancement by Chinese AI developers, which has reignited debates over how to manage different classes of AI models. By distinguishing between closed frontier models and open models, the White House appears to be adopting a bifurcated regulatory strategy that subjects proprietary high-capability systems to government scrutiny while leaving open-weight ecosystems largely untouched by this specific evaluation mechanism. However, because the definitions driving this distinction remain classified or internal, the practical boundary between regulated and unregulated development remains opaque to the broader AI community.
What's confirmed, what's disputed
- ConfirmedThe White House does not plan to publicly release its new framework for evaluating advanced AI models.
- ConfirmedThe framework defines a covered frontier model as closed-source with state-of-the-art capabilities and national security risks.
- ConfirmedOpen models are explicitly excluded from the framework's testing and evaluation requirements.
- ConfirmedThere is no clear public definition of what constitutes state-of-the-art capabilities or national security risk under the framework.
- ConfirmedThe European Union declined to comment on whether it qualifies as a trusted partner under the undisclosed framework.
- ConfirmedDuring the 30-day pre-release review, models must be stored in high-security environments with detailed access logs.
The strongest case each way
Keeping the framework private means companies, policymakers, researchers and U.S. allies outside the process will be left guessing how the administration plans to implement one of its key AI policies, undermining the voluntary nature of the regime.
The framework addresses national security risks and requires high-security environments with detailed access logs during review, implying that public disclosure could compromise the integrity of the evaluation process or reveal sensitive government assessment capabilities.
Times this happened before
- Vulnerabilities Equities Process (VEP) · 2024Government retained zero-days secretly despite criticism, leading to calls for legislative oversight and transparency reforms.
- Export Administration Regulations (EAR) AI Controls · 2024Secret licensing criteria created industry confusion and compliance delays until clarifications were issued months later.
What's at stake
Companies excluded from private briefings cannot prepare for compliance, risking delayed product releases during the 30-day review window. International allies face uncertainty about trusted partner status, potentially fragmenting transatlantic AI governance. Open model developers gain temporary exemption but face long-term regulatory ambiguity if definitions of 'frontier' expand. The secrecy undermines trust in voluntary regimes, potentially pushing stakeholders toward binding legislation or alternative standards bodies. Research communities lose visibility into government safety benchmarks, hindering independent validation of AI risk assessments.
Noise Level
The timeline
Axios reports framework secrecy plan
Publication reveals White House intends to keep AI evaluation criteria private based on three sources.
- 2 days ago
Private industry briefings held
White House conducted staff-level meetings with select companies to review the completed framework.
The full record
Sources & methodology
Every claim above traces to these primary items. How we score →
Where the sources disagree
In dispute Critics assert that secrecy prevents accountability and leaves stakeholders guessing about implementation.
Established It is established that the framework is not public and lacks clear definitions for key terms, but the administration's specific rationale for this opacity beyond general security is not fully detailed in available sources.
What's being under-reported
Coverage lacks perspective from civil society organizations and academic AI safety researchers who would be directly impacted by inability to audit or validate government evaluation standards. Their absence obscures the technical feasibility concerns of implementing secret benchmarks and the long-term effects on independent safety research ecosystems.
Who changed their mind, and why
- White HouseTransitioned from developing a voluntary framework to enforcing it through private, staff-level industry briefings without public documentation. (was: Implied commitment to voluntary safety standards through previous executive actions.)
- European UnionMaintained strategic ambiguity by declining to comment on trusted partner status rather than endorsing or criticizing the secretive approach.
The forecast
Excluded firms and allied governments will likely pressure the administration for partial disclosure because opacity threatens the legitimacy of voluntary safety pacts essential for international coordination.
Forecast, not fact — an editorial estimate we score when this resolves.
That's the complete picture as of — nothing more to know right now. We'll update this page the moment it changes.
Join the Discussion
Discuss this story
Community comments coming in a future update
Be the first to share your perspective. Subscribe to comment.