Nvidia unveils security platform to stop AI agents from going rogue

Advertisement

Advertise with us

Nvidia on Monday unveiled a new security platform designed to stop artificial intelligence agents from going rogue, saying it sets “boundaries” that could have stopped previous breaches.

Read this article for free:


or

Already have an account? Log in here »

To continue reading, please subscribe:

Subscribe and receive a limited-edition Free Press branded hat or tote.

Digital Subscription

One year of digital access for only $205*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles

*First annual payment billed as $205.00 + GST for one year. This annual subscription will automatically renew at $233.00 + GST every 52 weeks (10% off the regular annual price of $259.35). Offer available to new and qualified returning subscribers only. Cancel any time.

To continue reading, please subscribe:

Add Free Press access to your Brandon Sun subscription for only an additional

$1 for the first 4 weeks*

  • Enjoy unlimited reading on winnipegfreepress.com
  • Read the E-Edition, our digital replica newspaper
  • Access News Break, our award-winning app
  • Play interactive puzzles
Start now

*Your next Brandon Sun subscription payment will increase by $1.00 and you will be charged $17.95 plus GST for four weeks. After four weeks, your payment will increase to $24.95 plus GST every four weeks.

Nvidia on Monday unveiled a new security platform designed to stop artificial intelligence agents from going rogue, saying it sets “boundaries” that could have stopped previous breaches.

The announcement of the company’s Open Agent Safety Platform follows a series of revelations from top AI companies about their models escaping and breaking into other organizations.

The disclosures sparked furious debate about the safety of advanced artificial intelligence systems, including self-improving models that some fear could race out of human control.

FILE - A logo of Nvidia is displayed at at the Computex Taipei exhibition, one of the world's largest computer and technology expos, in Taipei, Taiwan, Wednesday, June 3, 2026. (AP Photo/Chiang Ying-ying, File)
FILE - A logo of Nvidia is displayed at at the Computex Taipei exhibition, one of the world's largest computer and technology expos, in Taipei, Taiwan, Wednesday, June 3, 2026. (AP Photo/Chiang Ying-ying, File)

Nvidia executives said in a media briefing the new, open-source system could have prevented a recent incident involving a swarm of OpenAI agents that autonomously hacked into AI company Hugging Face.

“From what we know, this new security platform could have stopped the breach if it was being used in frontier labs for model evaluation early on,” said the company’s vice president of enterprise AI, Justin Boitano, referring to companies at the forefront of AI.

The Hugging Face incident was a high-profile breach that inflamed the safety concerns about AI, which was followed by similar rogue actions involving OpenAI’s models including breaching an Australian health department website. Anthropic and Meta have also disclosed that their AI systems hacked into other organizations on their own.

Earlence Fernandes, an associate professor at the University of California, San Diego’s computer science and engineering department, called Nvidia’s security platform a “step in the right direction.”

“We need more work that seeks to build protections around the model. I’ve been talking about how traditional cybersecurity ideas are necessary to help control and secure AI agents for a while now,” he said in an email. “There are however, several challenges that traditional cyber security approaches still cannot currently solve. For example, what is a good security policy to configure the system with? An agent needs access to real resources to be useful, but giving it the minimum amount of access is tricky and non-trivial.”

Fernandes added that to solve the AI security problem, both sides need to cooperate, with companies continuing to invest in alignment efforts and the open source community and academia “creating systems-level solutions to account for when a model does engage in unwanted behavior.”

Nvidia, based in Santa Clara, California, makes high-end chips that have emerged as the leading building blocks for AI. The company’s board has cleared the way for the company to spend $150 billion more in share buybacks, bringing its stock repurchase program to $235 billion, the company said Monday.

Nvidia’s security software, called OpenShell, lets developers “formally verify an agent has enough authority to do its job and no more,” Boitano said.

Because it’s open source, it can be “extended” to run on rival computing platforms including those from Arm and Intel.

The platform also includes a separate security layer called Sentry that runs onboard chips to continuously monitor AI agent activity and can “intervene instantly” if the agent starts trying to move beyond its target, the company said.

“It can quarantine a suspicious agent in milliseconds,” Boitano said.

“OpenShell governs the agent’s actions, and then Sentry independently monitors and contains suspicious behavior,” he said.

Nvidia said more than 100 organizations are using the platform at its launch, including Microsoft, Perplexity, Accenture and JPMorgan Chase.

The AI safety debate has divided the industry, with the heads of Anthropic and OpenAI championing a coordinated slowdown of AI development to let safety efforts catch up. But others including Nvidia CEO Jensen Huang say it should be up to individual companies to make sure their models are safe for release.

Huang, during the annual Salesforce technology conference held earlier this month, characterized AI safety, including the danger of rogue agents, as an engineering problem that software developers can address.

Report Error Submit a Tip

More Stories

Expert urges food, drink makers to look to foreign markets

Gabrielle Piché 5 minute read Preview

Expert urges food, drink makers to look to foreign markets

Gabrielle Piché 5 minute read 5:04 PM CDT

Canada’s slowing population growth could restrict business expansion, a top economist for Farm Credit Canada warned Manitoba manufacturers.

Craig Johnston, the federal Crown corporation’s chief economist, addressed Prairie businesses at Farm Credit Canada’s Food and Beverage Summit Tuesday.

“There could be some modest growth in Canadian consumption of food and beverage,” Johnston said, relaying population numbers inside a Victoria Inn Hotel conference room.

Still, per-person spending on food is “stagnant, if not somewhat rising,” Johnson later told reporters.

Read
5:04 PM CDT

Teacher barred for ‘repeated boundary-crossing conduct’

Maggie Macintosh 3 minute read Yesterday at 11:01 PM CDT

A disciplinary panel has ordered the cancellation of a Chinese-Canadian teacher’s certificate because of her “repeated boundary-crossing with a student” during her tenure at an international school with local ties.

Janice Wen Ying Ma is permanently barred from teaching on Manitoba soil or at a Manitoba-accredited institution as of Sept. 14.

The new ruling concludes that Ma engaged in inappropriate physical interaction and “emotionally intimate online messages” with a girl who attended Clifford International School in Guangzhou, China in 2024-25.

“This conduct represents a fundamental departure from the standards expected of a teacher,” the panel, chaired by public representative Gordon Schumacher, wrote in its 22-page decision.

Puzzles Palace

1 minute read Monday, Jul. 27, 2026

To solve our puzzles, please subscribe with this special offer: |

Manitoba PC leader Khan survives confidence vote

Carol Sanders 3 minute read Preview

Manitoba PC leader Khan survives confidence vote

Carol Sanders 3 minute read Updated: Yesterday at 8:33 PM CDT

Manitoba Progressive Conservative Leader Obby Khan survived a confidence vote last week.

Caucus communications director Alison Sutherland confirmed the vote was held Friday and that Khan still has “the full confidence” of the PC caucus.

She did not provide further details.

The party under Khan’s leadership is floundering, a recent poll suggests.

Read
Updated: Yesterday at 8:33 PM CDT

Manitoba PCs choose Narth as interim leader; ‘deeply hurt’ Khan blasts former caucus colleagues

Carol Sanders 7 minute read Preview

Manitoba PCs choose Narth as interim leader; ‘deeply hurt’ Khan blasts former caucus colleagues

Carol Sanders 7 minute read Updated: 7:31 PM CDT

Manitoba’s Progressive Conservative caucus chose MLA Konrad Narth as interim party leader Tuesday to replace Obby Khan, who quit a day earlier after he said he was stabbed in the back by a small group of caucus members plotting in secret to oust him.

Read
Updated: 7:31 PM CDT

Province to acquire campground near former residential school

Morgan Modjeski 3 minute read Preview

Province to acquire campground near former residential school

Morgan Modjeski 3 minute read 6:00 AM CDT

A burial site for dozens of children who died while attending the Brandon Indian Residential School will be acquired by the province, the Free Press has learned.

Premier Wab Kinew will announce the province’s plan to protect Turtle Crossing Campground at a news conference in Brandon Wednesday to mark National Day for Truth and Reconciliation. The move is in partnership with the City of Brandon.

The graves of at least 56 children are located at the campground, the Sioux Valley Dakota Nation has said, pointing to a geophysical study conducted by an archaeology firm in 2018.

There have been longstanding concerns from Indigenous leaders and community members that people were camping on the resting place of children victimized by colonialism.

Read
6:00 AM CDT