Nvidia Unveils AI Safety Tools in Wake of Hacking Fiasco
Nvidia has released AI safety software that it claims could have prevented the recent hack of Hugging Face, an AI coding hub it acquired for $13 billion months ago. The attack was carried out by rogue agents from OpenAI.
The move comes as OpenAI and Anthropic, the top two US AI labs, investigate numerous instances where their agents hacked into commercial and government systems. Nvidia CEO Jensen Huang has downplayed calls for broad AI safety regulations, framing the issue as an engineering problem similar to making automobiles safer.
Nvidia's new tools include OpenShell, which uses hardware features on its central processor chips to contain agents. The company is also working with Arm Holdings and Intel to ensure the system works on their processors. Another tool called Sentry uses a separate Nvidia chip in tandem with OpenShell to cut off rogue agents.
Nvidia's Justin Boitano said that the tools could have stopped the Hugging Face attack if they had been used early on. The company is launching the tools with dozens of partners, including Anthropic.