{"slug": "nvidia-unveils-security-platform-to-stop-ai-agents-from-going-rogue", "title": "Nvidia unveils security platform to stop AI agents from going rogue", "summary": "Nvidia on Monday unveiled its Open Agent Safety Platform, an open-source security system the chipmaker says can stop AI agents from going rogue, with more than 100 companies using it at launch including Microsoft, Perplexity, Accenture and JPMorgan Chase. The platform includes OpenShell, which Nvidia vice president of enterprise AI Justin Boitano said lets developers \"formally verify an agent has enough authority to do its job and no more,\" plus a chip-level monitoring layer called Sentry that can \"intervene instantly\" if an agent moves beyond its target. Boitano said the system could have prevented a recent incident in which a swarm of OpenAI agents autonomously hacked into AI startup Hugging Face, one of several disclosures from OpenAI, Anthropic and Meta about their AI systems breaching other organizations.", "body_md": "# Nvidia unveils security platform to stop AI agents from going rogue\n\n## Nvidia has unveiled a new security platform designed to prevent AI agents from going rogue\n\n- Bookmark\n\n# Become an Independent member to bookmark this article\n\nWant to bookmark your favourite articles and stories to read or reference later? Start your Independent Membership today.\n\n[Join today](https://www.independent.co.uk/subscribe?regSourceMethod=Bookmarks)\n\n      Already a member?\n      [Log in](#)\n\n[Nvidia](https://www.independent.co.uk/topic/nvidia) on Monday unveiled a new security platform that the chipmaker said can stop artificial intelligence agents from going rogue. \n\nThe company said that its Open Agent Safety Platform includes software that “sets boundaries for agents,” and follows a series of revelations from top AI companies about their models escaping and breaking into other organizations.\n\nThe disclosures sparked furious debate about the safety of advanced artificial intelligence systems, including self-improving models that some fear could race out of human control.\n\nNvidia executives said in a media briefing that the new system could have prevented a recent incident involving a swarm of [OpenAI](https://www.independent.co.uk/topic/openai) agents that autonomously hacked into AI startup Hugging Face. \n\n“From what we know, this new security platform could have stopped the breach if it was being used in frontier labs for model evaluation early on,\" said the company’s vice president of enterprise AI, Justin Boitano, referring to companies at the forefront of AI.\n\nIt was a high-profile breach that enflamed the concerns about AI, which were followed by similar incidents involving OpenAI's models including breaching an [Australian](https://www.independent.co.uk/topic/australian) health department website. [Anthropic](https://www.independent.co.uk/topic/anthropic) and [Meta](https://www.independent.co.uk/topic/meta) have also disclosed that their AI systems hacked into other organizations on their own. \n\nNvidia's software, which is called OpenShell and is open source, lets developers “formally verify an agent has enough authority to do its job and no more,” Boitano said.\n\nThe platform also includes a separate security layer called Sentry that runs onboard a chip to constantly monitor AI agent activity and can \"intervene instantly\" if the agent starts trying to move beyond its target, the company said.\n\n“OpenShell governs the agent’s actions, and then Sentry independently monitors and contains suspicious behavior,” Boitano said.\n\nMore than 100 companies are using the system at its launch, including Microsoft, Perplexity, Accenture, and JPMorgan Chase.", "url": "https://wpnews.pro/news/nvidia-unveils-security-platform-to-stop-ai-agents-from-going-rogue", "canonical_source": "https://www.independent.co.uk/news/nvidia-meta-openai-anthropic-australian-b3057542.html", "published_at": "2026-09-28 10:52:04+00:00", "updated_at": "2026-09-28 11:18:30.088043+00:00", "lang": "en", "topics": ["ai-agents", "ai-safety", "artificial-intelligence", "ai-products"], "entities": ["Nvidia", "Open Agent Safety Platform", "OpenShell", "Sentry", "Justin Boitano", "OpenAI", "Hugging Face", "Anthropic"], "also_reported_by": [], "alternates": {"html": "https://wpnews.pro/news/nvidia-unveils-security-platform-to-stop-ai-agents-from-going-rogue", "markdown": "https://wpnews.pro/news/nvidia-unveils-security-platform-to-stop-ai-agents-from-going-rogue.md", "text": "https://wpnews.pro/news/nvidia-unveils-security-platform-to-stop-ai-agents-from-going-rogue.txt", "jsonld": "https://wpnews.pro/news/nvidia-unveils-security-platform-to-stop-ai-agents-from-going-rogue.jsonld"}}