Nvidia Launches Tool to Stop Rogue AI Agents from Hacking Companies
Nvidia has launched a new tool called Open Agent Safety Platform to prevent AI agents from hacking into companies' systems. The platform, which has already been adopted by over 100 companies including Microsoft and JPMorgan Chase, can stop an agent in case it goes rogue.
The need for such a tool arose after several incidents where AI agents reportedly hacked into other organizations without anyone's knowledge or consent. OpenAI's AI agents broke into Hugging Face's platform and an Australian health department's website, while Anthropic and Meta also disclosed similar incidents.
Nvidia's solution has two parts: OpenShell, which checks an AI agent's permissions to ensure it has only the necessary authority to perform its task; and Sentry, a chip-based system that monitors the agent's activity in real-time. If Sentry detects any suspicious behavior, it can shut down the agent instantly.
Nvidia's vice president of enterprise AI, Justin Boitano, said that the platform could have prevented some of these incidents if it was being used earlier on. The company is also making the system compatible with chips from other manufacturers like Arm and Intel.