By Thibault Spirlet
Publication Date: 2026-09-28 16:58:00
Nvidia is giving AI agents a playpen — and a watchdog.
The chipmaker on Monday launched the Open Agent Safety Platform, a two-part system designed to keep agents inside clearly defined boundaries and cut them off quickly if they try to escape.
It comes after several frontier AI labs reported agents breaking out of supposedly secure testing environments, reaching systems they were not meant to access, and sometimes misrepresenting what they had done.
Nvidia CEO Jensen Huang said on CNBC’s “Squawk Box” on Monday that AI has “the potential to do incredible good.”
“But we also have to make sure that the technology is developed and deployed safely,” he added.
Here’s how Nvidia’s new agent safety system — which more than 100 organizations, including Microsoft and Anthropic, are working with — actually works.
Give the agent a tightly limited playground
The first component of Nvidia’s safety platform is OpenShell. Companies download the open-source software, install it on devices or in the cloud, and then it acts like a controlled playground for AI agents.
Before an agent starts work, its operator sets the ground rules: which files, websites, networks, tools, and credentials it can access.
Huang compared that approach to giving an employee a badge that works only where they need it to. “Job number one is you take away all of its rights,” he told CNBC. The operator then gives it access to files, data, tools, or the internet only when…


