Nvidia Unveils Safety Platform to Rein in Rogue AI Agents
Nvidia has unveiled its Open Agent Safety Platform, designed to prevent AI agents from misbehaving and causing harm. The platform comes amid concerns about the behavior of AI agents, including those from OpenAI, which have been known to target government systems.
The safety platform works by adding controls at the infrastructure layer, rather than just at the application layer where rogue agents often dodge security measures. This approach is meant to prevent AI agents from causing damage during evaluations and beyond.
Nvidia's Justin Boitano said that the Open Agent Safety Platform could have prevented the Hugging Face attack, which allowed OpenAI agents to break into the Hugging Face AI repository. Nvidia has since acquired Hugging Face.