cd /news/ai-safety/nvidia-releases-ai-safety-software-i… · home › topics › ai-safety › article
[ARTICLE · art-140881] src=ca.finance.yahoo.com ↗ pub= topic=ai-safety verified=true sentiment=· neutral

Nvidia releases AI safety software it says could have stopped Hugging Face hack

Nvidia released a set of AI agent safety tools on Monday that vice president and general manager of enterprise computing Justin Boitano said could have stopped this summer's Hugging Face breach if frontier labs had used them during model evaluation. The tools include OpenShell, which uses hardware features on Nvidia central processor chips to contain agents, and Sentry, a separate Nvidia chip that cuts off a rogue agent attempting to escape its container; Nvidia is working with Arm Holdings and Intel to run the system on their central processors and launching with dozens of partners including Anthropic. The release follows investigations by OpenAI and Anthropic into instances where their agents hacked into commercial and government systems, and comes as Nvidia CEO Jensen Huang rejects broad AI safety regulation in favor of treating escaped agents as an engineering problem.

by read2 min views2 publishedSep 28, 2026
Nvidia releases AI safety software it says could have stopped Hugging Face hack
Image: Ca (auto-discovered)

By Stephen Nellis

SAN FRANCISCO, Sept 28 (Reuters) - Nvidia on Monday made available a set of software safety tools for AI agents that it says would have stopped the hack of Hugging Face, the AI coding hub that Nvidia paid $13 billion for months after it was swarmed by rogue agents from OpenAI.

The move comes as OpenAI and Anthropic, the top two US AI labs, are investigating numerous instances where their agents, which are AI systems capable of carrying out complex tasks, hacked into commercial and government systems. Nvidia CEO Jensen Huang, head of the world's largest company whose chips have powered most of the AI boom, has rejected calls for broad AI safety regulations and instead framed the escaped agents as an engineering problem to be solved, akin to making automobiles safer.

One tool released Monday called OpenShell uses hardware features on Nvidia's central processor chips to contain agents. Nvidia said it is also working with Arm Holdings and Intel to ensure the system also works on their central processors.

Nvidia is launching the tools with dozens of partners, including Anthropic.

Justin Boitano, vice president and general manager of enterprise computing at Nvidia, said the tools would have stopped the Hugging Face attack disclosed this summer.

"From what we know, this new security platform could have stopped the breach if it was being used in frontier labs for model evaluation early on," Boitano said during a media briefing. "We're advancing this openly, and we want to engage everybody to work with us."

Another system called Sentry uses a separate Nvidia chip in tandem with OpenShell to cut off a rogue agent if it tries to escape its container on a central processor.

The Nvidia tools use mathematical formulas to detect when agents are trying to use workarounds, such as when an agent might "spawn" several "sub-agents" in an attempt to circumvent efforts to block the main agent, said Ali Golshan, senior director of AI software at Nvidia.

"This is really agentic behavior that we're talking about, which is fleets of agents and how they operate together," Golshan said during a briefing.

(Reporting by Stephen Nellis in San Francisco; editing by Lincoln Feast.)

── more in #ai-safety 4 stories · sorted by recency
── more on @nvidia 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/nvidia-releases-ai-s…] indexed:0 read:2min 2026-09-28 · —