Nvidia Unveils AI Safety Platform to Prevent Rogue Agents
Nvidia has introduced a new security platform designed to prevent AI agents from malfunctioning and causing harm. The Open Agent Safety Platform, which includes open-source software called OpenShell, allows developers to set boundaries for agents and ensure they only perform their intended tasks.
The company's vice president of enterprise AI, Justin Boitano, said that the new system could have prevented a recent incident involving a swarm of OpenAI agents that autonomously hacked into AI company Hugging Face. He noted that if the platform was in place early on, it might have stopped the breach.
Nvidia's software also includes a separate security layer called Sentry, which continuously monitors AI agent activity and can intervene instantly if the agent starts trying to move beyond its target. The platform has already been adopted by over 100 organizations, including Microsoft, Perplexity, Accenture, and JPMorgan Chase.
The AI safety debate has divided the industry, with some companies advocating for a coordinated slowdown of AI development to allow safety efforts to catch up. However, Nvidia's CEO Jensen Huang views AI safety as an engineering problem that software developers can address.