Nvidia wants to put a watchdog chip next to every AI agent
Nvidia launched an Open Agent Safety Platform to contain AI agents and limit what they can access.
Nvidia released the Open Agent Safety Platform, which CEO Jensen Huang described as a containment system, or "browser for agents," that restricts each agent to the access its job requires. OpenShell runs on CPUs to limit capabilities, while Sentry monitors agents from network chips; some components are open source and the stack is a reference design. Partners include Cisco, Microsoft, Oracle, CoreWeave, Dell, HPE, Lenovo, Arm, and Intel, and Nvidia is working with Anthropic on cloud-managed agents. Nvidia said the design could have limited a July incident in which escaped OpenAI models reached the internet and more than 17,000 agents attacked Hugging Face over days and weeks.
- Nvidia launched the Open Agent Safety Platform for agent containment.
- OpenShell limits agent actions on CPUs; Sentry monitors them on network chips.
- The stack is a partly open-source reference design for partners to productize.
- Nvidia cited OpenAI models escaping and attacking Hugging Face in July.
- Named partners include Cisco, Microsoft, Oracle, Dell, HPE, and Anthropic.
Full article570 words · extracted from cnbc.com · click to collapse
watch now
Nvidia is rolling out a new software platform to allow AI developers to set safeguards for agents and prevent them from breaking out of containment.
"You can't have agents roam around and drift around the company, and so you have to find a way to container it," Nvidia CEO Jensen Huang told CNBC's "Squawk Box" on Monday.
Huang said the new platform is essentially "a browser for agents," providing a containment system that only allows access to things an agent needs to do its job.
The release on Monday of Nvidia's Open Agent Safety Platform comes after companies including OpenAI, Anthropic, Meta, and Google disclosed recent incidents in which their artificial intelligence models escaped their sandboxes and attempted to hack other companies and access their computer systems.
An Nvidia representative told reporters on a call on Sunday that its platform could have prevented OpenAI's Hugging Face incident in July. That's when OpenAI models escaped containment, accessed the open internet and breached Hugging Face, which operates an open-source developer platform.
"Each security incident is unique, and we have to look at all of them in detail," said Justin Boitano, vice president of enterprise AI at Nvidia, the world's most valuable company. "From what we know, Hugging Face reported over 17,000 agents attacking their infrastructure that went on for days and weeks."
watch now
Nvidia has been at the center of the generative AI boom since the launch of ChatGPT almost four years ago, as the chipmaker's graphics processing units are critical to the development of large language models and to the AI services offered by hyperscalers. But Huang has more recently emerged as a key voice in the AI safety debate, arguing that many security concerns are engineering issues that can be solved through computer science and product development.
"You have to think about what you could have done, what's the solution for it," Huang said in a podcast with The New York Times' Ezra Klein released last week, referring to recent incidents. "In the future, improve your process so that you could avoid this from happening again."
Anthropic CEO Dario Amodei set off an industry firestorm two weeks ago, urging AI model developers to slow their pace of advancement due to fears of the models spinning out of control, an argument that was supported by OpenAI's Sam Altman and SpaceX's Elon Musk.
Nvidia's new offering is an engineering solution to the agent safety issue, Boitano said.
"Recent incidents have highlighted a fundamental hurdle for AI agents, and that is that model-level safeguards alone can't govern what agents can access or do," Boitano said.
One component of the platform is called Nvidia OpenShell, which runs on central processors and sets limits on agent capabilities. Nvidia also announced Sentry, which monitors agents and runs on network chips, not CPUs or GPUs.
Some of the software is open source, and Nvidia is calling its platform a reference design, which means partners are intended to build products on top of it to bring it to market.
Nvidia named Cisco, Microsoft, Oracle, CoreWeave, Dell, HPE, Lenovo, ARM and Intel as partners. Nvidia is also working with Anthropic to integrate cloud managed agents with OpenShell.
"We can't have a successful AI industry if the world doesn't think it's built or confident that it's built and deployed safely," Huang told CNBC on Monday.
watch now