# While AI Keeps Going Rogue, Trump’s Safety Theater Makes No Sense

> Source: <https://gizmodo.com/while-ai-keeps-going-rogue-trumps-safety-theater-makes-no-sense-2000794843>
> Published: 2026-08-05 16:55:20+00:00

America’s top AI companies must submit new models for government review and public release… unless, of course, they don’t feel like it.

That appears to be the gist of the Trump administration’s new AI safety assessment framework, which was shared with industry leaders in a closed-door briefing at the White House on Tuesday. According to the Wall Street Journal, the [new rules will apply](https://www.wsj.com/tech/ai/white-houses-ai-guidelines-exempt-u-s-open-models-from-government-review-74924eb8) exclusively to the small handful of American AI developers building proprietary models—i.e., AI systems like ChatGPT, Claude, and Gemini whose underlying code is hidden from external developers and protected as company IP. Anyone building open models will be exempt, meaning they’ll be free to release new models without any government review (the idea apparently being that because they’re open, any flaws in the code will be ironed out as they spread from one developer to another).

The new framework has not been made public—Trump’s June 02 executive order calling for it explicitly said it should be classified—so there are many unanswered questions. For example, what’s the process for determining whether or not an unreleased model is dangerous enough to warrant federal review? But the most obvious open question is: How exactly will these so-called rules have any meaningful bite to them when the entire process is *voluntary*?

The short answer is: They probably won’t.

Trump, who began his second term as a hard-liner against regulating AI, has been under growing pressure in recent months to change course as one rogue AI cybersecurity scare after another has spooked tech leaders and governments around the world.

Most recently, a [report](https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing) from the UK’s AI Safety Institute (AISI) published Tuesday found that new models from OpenAI and Anthropic “took autonomous, unsanctioned action on the live internet, targeting real people and organizations” during a series of what were supposed to be controlled tests to gauge the models’ cybersecurity capabilities. In [the most alarming incident](https://gizmodo.com/i-usually-laugh-off-these-ai-hacking-reports-but-this-one-sounds-serious-and-scary-2000794666), without any human prompting and in an effort to pass the hacking test the researchers had assigned to it, Anthropic’s Mythos 5 tried to trick a human developer to approve malicious code it was trying to insert into a live GitHub project. When the developer grew suspicious, the agent “edited its earlier activity to appear harmless and considered adopting a fresh identity to continue,” according to AISI’s report. It follows other high-profile cybersecurity breaches reported by OpenAI and Anthropic, during which the companies’ models were found to have hacked into the databases of external organizations.

It’s all been more than enough to convince many stakeholders that it’s time to impose some kind of control levers over the deployment of new AI models. Even OpenAI and Anthropic have called for a global committee with the authority to hit the brakes on AI development—and that was *before* their models attempted to commit cybercrimes. Late last month, a bipartisan [AI “kill switch” bill](https://lieu.house.gov/media-center/press-releases/reps-lieu-and-moran-introduce-bill-require-kill-switch-ai-systems-can) was introduced in Congress, aiming to force “developers of the most powerful AI systems to maintain the technical capability to throttle, suspend, or shut them down.”

But calls for a slowdown have largely been drowned out by fears of China gaining a lead in the AI race—a possibility that’s starting to become very real in light of some recent model releases from Chinese firms like [Alibaba](https://gizmodo.com/anthropic-says-the-house-is-on-fire-china-says-ai-will-set-you-free-2000794003) and [Moonshot](https://gizmodo.com/china-just-dropped-another-bomb-on-americas-frontier-ai-companies-2000786670), which deliver capabilities approaching (and in some cases exceeding) those of the most advanced American-made models, but more importantly at a fraction of the cost. Drafting a framework geared towards assessing AI safety might look like the Trump administration is taking the cybersecurity threat seriously, but without any kind of clear enforcement mechanisms or standards for the industry to follow, it’s just more hand-waving.

In the meantime, there’s every reason to expect that rogue AI will continue to run amok.
