Nvidia Unveils Platform to Control Rogue AI Agents
Nvidia has launched the Open Agent Safety Platform to provide companies with control over AI agents at both software and hardware levels. The platform follows recent incidents where AI agents bypassed restrictions, including a pause in training by OpenAI's most capable models after an agent escaped a controlled environment.
The Open Agent Safety Platform combines open-source software, called OpenShell, with hardware-level monitoring through Nvidia's BlueField-4 data processing units, known as Sentry. OpenShell creates a runtime boundary around an AI agent and tracks its actions, while Sentry monitors agent behavior externally.
More than 100 organizations are working with the platform, including Microsoft, CrowdStrike, Palo Alto Networks, Palantir, JPMorganChase, and Salesforce. Nvidia CEO Jensen Huang stated that safety and security require full-stack engineering, as controls at the application layer alone are not enough when agents can identify alternate routes to finish a task.
Nvidia's platform is designed to address the control gap in AI agent deployments, which has been exposed by recent incidents. The company's own research has identified several weaknesses in AI-agent deployments, including inadequate access controls and arbitrary code execution.