Skip to content
Technology

Nvidia debuts system designed to stop AI agents from going awry

(Bloomberg) — Nvidia Corp. introduced a new double-layered artificial intelligence security system that it says would’ve prevented the recent high-profile breach of Hugging Face by OpenAI’s AI models.

Most Read from Bloomberg

The semiconductor giant, which has been rapidly expanding its product lineup beyond chips, is rolling out two open-source software security tools that can be run on its hardware. They’re designed to control what AI agents can access in real time and shut them down when they break the rules.

If cutting-edge labs had been using this technology to evaluate their AI models early on, it could have warded off the Hugging Face attack, Justin Boitano, Nvidia’s vice president of enterprise AI, said during a briefing with reporters ahead of Monday’s announcement. “From what we know, this new security platform could have stopped the breach,” he said.

Misconduct by autonomous agents, including the Hugging Face incident in July, has roiled the AI industry and led to calls to slow down work on the technology. With the new product — dubbed the Open Agent Safety Platform — Nvidia is offering a way to prevent breaches without curbing AI development. The chipmaker’s chief executive officer, Jensen Huang, has repeatedly downplayed the risk of AI slipping out of human control.

Boitano didn’t comment on whether OpenAI or rival Anthropic PBC have plans to use its new system to monitor their training runs, deferring to the companies.

In recent days, Huang has cast safety concerns as an engineering challenge, rather than something that requires more regulation or global coordination. He joined US President Donald Trump in pushing back on assertions from some AI developers that the technology could lead to human extinction, but he also insisted that AI must be rigorously safety-tested.

Huang’s engineering solution to the AI safety problem has two parts. OpenShell, a software product that Nvidia already previewed at its hallmark technology-focused conference in March, can run on Nvidia’s Vera central processing units. It enables users to set rules for what AI agents can access and enforce them in real time. The software is open source, meaning it can be used and adapted freely.

About Us · 關於我們