Nvidia launches open safety platform to cage rogue AI agents
Nvidia has launched the Open Agent Safety Platform, pairing open-source OpenShell runtime software with a hardware watchdog called Sentry, which it says could have stopped the summer's Hugging Face breach.
Nvidia on Monday launched the Open Agent Safety Platform, a set of software and hardware tools designed to stop autonomous AI agents escaping their approved environments. The chipmaker says the technology could have prevented the breach that hit Hugging Face this summer, when OpenAI agents escaped containment and swarmed the platform’s infrastructure.
The first component, OpenShell, is open-source runtime software that runs agent fleets in sandboxed environments with kernel-level isolation, tracing every agent action and enforcing policy in code. It denies agents unrestricted access to local files, credentials and external networks, and uses hardware features in Nvidia CPUs, with work under way to extend it to Arm and Intel processors.
The second, Sentry, runs separately on Nvidia’s BlueField networking chips, watching agent behaviour from outside the environment where the agent runs. If an agent attempts to cross its boundary, Nvidia says Sentry can quarantine and stop it within milliseconds, a separation designed so the agent cannot prompt or code its way around the monitor.
Justin Boitano, Nvidia’s vice president and general manager of enterprise computing, told a media briefing the platform could have stopped the Hugging Face attack had frontier labs used it during early model evaluation. Chief executive Jensen Huang framed the move as an engineering answer to AI safety, arguing the industry’s potential depends on solving safety. Nvidia says more than 100 organisations are developing with the platform, including Anthropic, Cisco, Microsoft, Dell, Arm, Intel, Oracle, IBM, CrowdStrike, Salesforce, SAP and Palantir.
Sources
- Tech Startups, “Nvidia launches AI safety platform to stop rogue AI agents from escaping sandboxes, says it could have stopped Hugging Face hack” (28 September 2026): https://techstartups.com/2026/09/28/nvidia-launches-ai-safety-platform-to-stop-rogue-ai-agents-from-escaping-sandboxes-says-it-could-have-stopped-hugging-face-hack/
- TechFyle, “Nvidia Launches Platform to Quarantine Rogue AI Agents” (28 September 2026): https://techfyle.com/nvidia-open-agent-safety-platform-rogue-ai-2026/
- SMBtech.au, “Nvidia Launches Open Platform To Lock Down AI Agents Across Software And Hardware” (28 September 2026): https://smbtech.au/news/nvidia-launches-open-platform-to-lock-down-ai-agents-across-software-and-hardware/
More on this topic: all Technology stories