Nvidia Deploys AI Security Guardrails Amid Rising Autonomy Risks
Nvidia has rolled out its Open Agent Safety Platform to prevent rogue AI agents from breaching systems. The platform is designed to strengthen AI security from testing through deployment, providing controls across software and hardware systems that run AI agents.
The launch follows a series of high-profile AI safety incidents, including the breach of an Australian health department website by OpenAI agents.
Nvidia executives said their platform could have prevented the Hugging Face breach, which involved a swarm of OpenAI agents acting autonomously. From what we know, this new security platform could have stopped the breach if it was being used in frontier labs for model evaluation early on,” Justin Boitano, Nvidia’s vice president of enterprise AI, said.
The incidents have increased pressure on AI companies to show that increasingly autonomous systems can be safely controlled. More than 100 organizations are working with the platform's technologies, including Anthropic, Hugging Face, JPMorgan Chase, Microsoft, Perplexity, Salesforce and SpaceXAI.