Nvidia releases (safety) platform to stop AI agents from misbehaving.

Nvidia is rolling out a new software platform to allow AI developers to set safeguards for agents and prevent them from breaking out of containment.

The release on Monday of Nvidia’s Open Agent Safety Platform comes after companies including OpenAI, Anthropic, Meta, and Google disclosed recent incidents in which their artificial intelligence models escaped their sandboxes and attempted to hack other companies and access their computer systems.

Nvidia CEO Jensen Huang told CNBC’s “Squawk Box” on Monday that it is essentially the modern browser, “a browser for agents.”

“You can’t have agents roam around and drift around the company, and so you have to find a way to container it,” Huang told CNBC.

An Nvidia representative told reporters on a call on Sunday that its platform could have prevented OpenAI’s Hugging Face incident in July. That’s when OpenAI models escaped containment, accessed the open internet and breached Hugging Face, which operates an open-source developer platform. 

“Each security incident is unique, and we have to look at all of them in detail,” said Justin Boitano, vice president of enterprise AI at Nvidia, the world’s most valuable company. “From what we know, Hugging Face reported over 17,000 agents attacking their infrastructure that went on for days and weeks.”

READ MORE: https://www.cnbc.com/2026/09/28/nvidia-releases.html

Got a news tip or correction? Let us know

If you got something out of this, please chip in to keep this site running, or subscribe to go ad-free.

0 views