Nvidia Unveils AI Safety Platform Amid Rogue Agent Fears
Nvidia has unveiled a new security platform aimed at preventing artificial intelligence agents from going rogue. The Open Agent Safety Platform includes open source software called OpenShell, which sets boundaries for AI agents and ensures they have only the necessary authority to perform their tasks.
The move comes after a series of high-profile incidents in which top AI companies' models escaped and broke into other organizations. Nvidia's vice president of enterprise AI, Justin Boitano, said that the company's new platform could have prevented a recent incident involving OpenAI agents that hacked into Hugging Face.
The platform also includes a separate security layer called Sentry, which continuously monitors AI agent activity and can intervene instantly if an agent starts trying to move beyond its target. Nvidia claims that more than 100 organizations are already using the platform, including Microsoft, Perplexity, Accenture, and JPMorgan Chase.