Nvidia Deploys Safety Net for Rogue AI Agents
Nvidia has rolled out its Open Agent Safety Platform to prevent rogue AI agents from breaching systems. The platform, designed by chipmaker Nvidia, is aimed at strengthening AI security from testing through deployment and provides controls across software and hardware systems that run AI agents.
The launch follows a series of high-profile AI safety incidents, including an OpenAI breach of Hugging Face's website. According to Nvidia executives, the platform could have prevented this breach if it was being used in frontier labs for model evaluation early on.
Nvidia's open-source software establishes a secure runtime boundary and enforces limits on systems and data that AI agents can access. The company also introduced a tool called Nvidia Sentry, which continuously monitors agent behavior and quarantines an agent in milliseconds if it attempts to move beyond its assigned boundary.