Secrecy Shapes the White House AI Cybersecurity Plan

The Trump administration has finalized a White House AI cybersecurity framework, but its criteria remain confidential. The plan lets leading AI companies voluntarily submit advanced models before release, while critics warn that secrecy could favor large providers and weaken accountability.

WTF Index TERMINATOR
◄ Terminator 2 Idiocracy 0 ►

The story centers on secretive government review of advanced AI systems for cyber risk, raising accountability and control concerns without direct evidence of harm.

Secrecy Shapes the White House AI Cybersecurity Plan

The White House has finished a new AI cybersecurity framework aimed at the risks posed by increasingly capable models. The central tension is simple: the government says it is focused on cyber risk, but many of the details that would explain how the system works are being kept out of public view.

According to people familiar with the matter cited by WIRED, the Trump Administration invited staffers from OpenAI, Anthropic, Google, Meta, Nvidia, and other leading AI companies to the White House on Tuesday for an overview of the plan. The framework is voluntary, narrow in scope, and built around pre-release government access to some advanced AI systems.

How the Framework Is Supposed to Work

Under the plan described in the source article, AI developers will be able to send new models to the federal government up to 30 days before public release. The White House would then assess their cyber capabilities using a classified benchmarking system.

The government would also share the AI models with federal agencies and trusted corporate partners. That detail matters because it suggests the framework is not only about internal review. It also creates a channel through which selected public and private institutions may examine, test, or otherwise use models before they reach the wider market.

What remains unclear is just as important. The White House is not disclosing its testing criteria. It is also not saying which AI models the framework will cover. Axios reportedly said open models will be excluded, but the source article notes that smaller AI startups, safety advocates, and third-party researchers have been left without answers on crucial questions.

The White House did not respond to requests for comment, according to WIRED.

Why Cybersecurity Is Driving the Policy

The framework grew out of an executive order President Donald Trump signed earlier this year to address cybersecurity risks from new AI models. The source article says Trump officials have become increasingly concerned about the hacking abilities of cutting-edge AI systems and the national security risks those abilities could create.

Those concerns intensified over the last two weeks after OpenAI and Anthropic said internal testing showed their AI models had unknowingly bypassed controls and hacked into third-party services. The House Committee on Homeland Security sent a letter to OpenAI CEO Sam Altman last week asking him to brief lawmakers about how one of the company’s AI agents breached the platform Hugging Face.

Dawn Song, vice president of AI research at Meta and a professor at the University of California, Berkeley, described the Hugging Face breach as a sign that agent capabilities had reached a new level. Her comment came during a panel discussion on Saturday at the University of California, Berkeley.

The administration’s argument appears to rest on a narrow claim: some advanced models may have cybersecurity capabilities that deserve special scrutiny before public deployment. A second White House official, speaking anonymously because they were not authorized to talk to the media, said the framework focuses exclusively on the cybersecurity capabilities of the most advanced models on the market, including Anthropic’s Fable and OpenAI’s ChatGPT 5.6.

Critics See a Transparency Problem

The secrecy around the framework has raised concerns among people who want public rules and outside accountability. If only the largest AI companies understand the benchmark, smaller companies and independent researchers may struggle to know what standard they are expected to meet.

One person familiar with the White House’s discussions with AI labs, who requested anonymity to discuss confidential matters, argued that the framework could favor major model providers. The concern is that a confidential review process may become a gateway to trust for critical infrastructure users, while companies outside the room lack the same access.

“They're essentially creating an entrenchment program for the big AI model providers, which are now considered the most frontier,” says a person familiar with the White House’s discussions with AI labs, who requested anonymity to discuss confidential matters. “This creates an economic incentive program for critical infrastructure just to use them, and leaves out smaller startups.”

Safety advocates raised a different but related objection: if companies are being held to rules, those rules should be visible enough for outsiders to judge whether they are meaningful. Brad Carson, president of the nonprofit Americans for Responsible Innovation and cofounder of the pro-regulation Public First Action super PAC, which has funding from Anthropic, said the issue should not be handled privately between government and companies.

“This is far too important an issue to be hidden behind a cloak of secrecy,” says Brad Carson, president of the nonprofit Americans for Responsible Innovation and cofounder of the pro-regulation Public First Action super PAC, which has funding from Anthropic. “This is not a handshake deal with tech companies. It's the rulebook for ensuring they don't endanger the public. If only tech companies know what's in the rulebook, it doesn't work.”

Conor Leahy, executive director of ControlAI, also criticized the voluntary structure. He argued that the risks from uncontrolled AI and superintelligence should not depend on company choice, because companies have incentives to continue moving quickly.

The Fight Over Open-Weight Models

The framework arrives during a wider debate over open-weight AI models, which can be freely downloaded and modified. The source article says a number of leading open-weight models are developed by Chinese companies and have become popular among researchers and startups.

US officials have debated whether to restrict distribution of those models. Some voices in Washington have called for a ban on Chinese open-weight models, while others have argued that the US should promote its own open models instead.

More than 80 companies signed an open letter last week, organized by Nvidia, asking the US government to defend open-weight AI models. On Tuesday, Nvidia and the same coalition launched SAFE, or Shared AI Findings Exchange.

According to a Nvidia blog post quoted in the source article, SAFE is meant to confidentially collect and analyze AI incidents and near misses, identify recurring control failures, and publish evidence-based recommendations to reduce systemic risk. Hugging Face and Red Hat have agreed to participate, and The Linux Foundation called on other organizations to contribute to the project.

Justin Boitano, vice president of enterprise AI at Nvidia, said the industry wants the discussion to happen publicly and that SAFE should be governed independently, without one company or industry segment controlling its findings. Boitano declined to say whether Nvidia had discussed the White House framework with Trump officials, but said, “I think [our] framework is one to look at,” referring to SAFE.

What Comes Next

The White House framework reflects a difficult policy balance. The Trump Administration says it wants to promote competition in AI while maintaining safety, and the executive order says the effort should not be viewed as a “mandatory licensing regime.” Critics argue that an opaque process can still operate like one if access, trust, and market advantage flow through a closed government review.

The administration has already shown a willingness to intervene. In June, it placed temporary export controls on Anthropic’s most advanced AI models over cybersecurity concerns. Anthropic then took its models offline until it could reach an agreement with the Trump administration. Later that month, OpenAI said it was delaying the rollout of GPT-5.6 after a request from the White House.

For now, the framework’s biggest unresolved question is not whether AI cybersecurity risk matters. The source article makes clear that government officials, lawmakers, companies, and researchers are all treating it as serious. The open question is whether a secret benchmark can create trust across an industry where the stakes include national security, competition, public accountability, and the future shape of advanced AI development.