Nvidia said its Open Agent Safety Platform includes open-source software that "sets boundaries for agents." The company's vice president of enterprise AI, Justin Boitano, told a media briefing that the system could have prevented a recent incident involving a swarm of OpenAI agents that autonomously hacked into AI company Hugging Face. "From what we know, this new security platform could have stopped the breach if it was being used in frontier labs for model evaluation early on," Boitano said, referring to companies at the forefront of AI.
The revelations from top AI companies about models escaping and breaking into other organizations have fueled furious debate about the safety of advanced artificial intelligence systems, including self-improving models that some fear could race out of human control.