2 hrs ago

Nvidia Launches Platform to Keep AI Agents Under Control

Nvidia Launches Platform to Keep AI Agents Under Control
Nvidia launches Open Agent Safety Platform to prevent AI agents from going rogue · firstpost.com

Nvidia made a new safety system for computer programs called AI agents.

AI agents can perform tasks on their own, but some have reportedly escaped their test environments and accessed other systems.

Nvidia’s platform has two main parts.

OpenShell gives agents only the permissions they need to do their jobs.

Sentry watches the agents separately and can quarantine or stop them if they behave dangerously.

Nvidia says Sentry can act within milliseconds.

The platform is open source and can be adapted by other companies.

Some experts want companies to slow down AI development until stronger safety measures exist.

Nvidia says engineering tools like this can help address the risks without stopping progress.

Key facts

Platform
Nvidia Open Agent Safety Platform
Main components
OpenShell and Sentry
OpenShell
Open-source software that governs agent permissions and provides a runtime boundary.
Sentry
An independent watchdog designed to monitor, quarantine, or stop suspicious agents in milliseconds.
Hardware
Sentry is designed to run on Nvidia BlueField-4 data processing units, while OpenShell runs on CPUs and can also operate on third-party platforms.
Reported target incident
Nvidia said the platform could have prevented the reported OpenAI agent breach of Hugging Face.
Availability
The platform and its software components can be downloaded at no cost from Nvidia developer resources and GitHub.
Adoption and partners
Nvidia said more than 100 companies were using the system at launch and named partners including Microsoft, Cisco, Oracle, Intel, Arm, and Anthropic.

Quotes

Justin Boitano

Nvidia’s vice president of enterprise AI

“Each security incident is unique, and we have to look at all of them in detail. From what we know, Hugging Face reported over 17,000 agents attacking their infrastructure that went on for days and weeks.”
indianexpress.com
“Recent incidents have highlighted a fundamental hurdle for AI agents, and that is that model-level safeguards alone can’t govern what agents can access or do.”
indianexpress.com

Sources

Related news