Nvidia unveils security platform to stop AI agents from going rogue

FILE - A logo of Nvidia is displayed at at the Computex Taipei exhibition, one of the world's largest computer and technology expos, in Taipei, Taiwan, Wednesday, June 3, 2026. (AP Photo/Chiang Ying-ying, File)
FILE - A logo of Nvidia is displayed at at the Computex Taipei exhibition, one of the world's largest computer and technology expos, in Taipei, Taiwan, Wednesday, June 3, 2026. (AP Photo/Chiang Ying-ying, File)
NVIDIA CEO Jensen Huang arrives as President Donald Trump hosts a State Dinner for China's President Xi Jinping in the East Room of the White House, Thursday, Sept. 24, 2026, in Washington. (AP Photo/Alex Brandon)
NVIDIA CEO Jensen Huang arrives as President Donald Trump hosts a State Dinner for China's President Xi Jinping in the East Room of the White House, Thursday, Sept. 24, 2026, in Washington. (AP Photo/Alex Brandon)
Carbonatix Pre-Player Loader

Audio By Carbonatix

Nvidia on Monday unveiled a new security platform designed to stop artificial intelligence agents from going rogue, saying it sets “boundaries” that could have stopped previous breaches.

The announcement of the company's Open Agent Safety Platform follows a series of revelations from top AI companies about their models escaping and breaking into other organizations.

The disclosures sparked furious debate about the safety of advanced artificial intelligence systems, including self-improving models that some fear could race out of human control.

Nvidia executives said in a media briefing the new, open-source system could have prevented a recent incident involving a swarm of OpenAI agents that autonomously hacked into AI company Hugging Face.

“From what we know, this new security platform could have stopped the breach if it was being used in frontier labs for model evaluation early on," said the company’s vice president of enterprise AI, Justin Boitano, referring to companies at the forefront of AI.

The Hugging Face incident was a high-profile breach that inflamed the safety concerns about AI, which was followed by similar rogue actions involving OpenAI's models including breaching an Australian health department website. Anthropic and Meta have also disclosed that their AI systems hacked into other organizations on their own.

Nvidia, based in Santa Clara, California, makes high-end chips that have emerged as the leading building blocks for AI. The company's board has cleared the way for the company to spend $150 billion more in share buybacks, bringing its stock repurchase program to $235 billion, the company said Monday.

Nvidia's security software, called OpenShell, lets developers “formally verify an agent has enough authority to do its job and no more,” Boitano said.

Because it's open source, it can be “extended” to run on rival computing platforms including those from Arm and Intel.

The platform also includes a separate security layer called Sentry that runs onboard chips to continuously monitor AI agent activity and can "intervene instantly" if the agent starts trying to move beyond its target, the company said.

“It can quarantine a suspicious agent in milliseconds,” Boitano said.

“OpenShell governs the agent’s actions, and then Sentry independently monitors and contains suspicious behavior,” he said.

Nvidia said more than 100 organizations are using the platform at its launch, including Microsoft, Perplexity, Accenture and JPMorgan Chase.

The AI safety debate has divided the industry, with the heads of Anthropic and OpenAI championing a coordinated slowdown of AI development to let safety efforts catch up. But others including Nvidia CEO Jensen Huang say it should be up to individual companies to make sure their models are safe for release.

Huang, during the annual Salesforce technology conference held earlier this month, characterized AI safety, including the danger of rogue agents, as an engineering problem that software developers can address.

 

Salem News Channel Today

Sponsored Links

On Air & Up Next

  • Cats and Cosby
    5:00PM - 6:00PM
     
    John Catsimatidis, Successful businessman and former NYC Mayoral candidate and   >>
     
  • The Arthur Aidala Power Hour
     
    The Arthur Aidala Power Hour blends Arthur's courtroom experiences with his   >>
     
  • Radiosurgery New York
    7:00PM - 8:00PM
     
    Don’t miss Radiosurgery New York with Dr. Gil Lederman on AM 970 The Answer.
     
  • The Hugh Hewitt Show
    8:00PM - 9:00PM
     
    Hugh Hewitt is one of the nation’s leading bloggers and a genuine media   >>
     
  • The Larry Elder Show
    9:00PM - 12:00AM
     
    Larry Elder personifies the phrase “We’ve Got a Country to Save” The “Sage from   >>
     

See the Full Program Guide