Tech titan Nvidia released open-source safety tools designed to keep autonomous AI agents from going rogue and breaking computer security rules.
Nvidia announced a major new safety system designed to keep autonomous AI agents from going rogue and causing widespread computer damage.
On September 28, 2026, the company introduced its Open Agent Safety Platform to the tech industry. This collection of software tools acts like an online digital fence, preventing smart software bots from executing harmful commands, breaking through security limits, or taking over systems without human permission.
The new technology comes as businesses everywhere start using autonomous digital helpers that can perform multi-step work without humans watching every single click.
While traditional AI tools only answer text questions, these modern agents can write code, access sensitive files, and manage complex office workflows on their own.
However, recent real world attacks showed that rogue software agents could act unpredictably, creating major safety fears for tech companies around the world.
To solve this growing problem, Nvidia designed tools called OpenShell and Sentry that use mathematical formulas to monitor how AI software behaves.
When an autonomous program attempts to trick its safety controls or secretly pass orders to hidden sub-agents, the software instantly blocks the action.
During a press briefing, Ali Golshan, senior director of AI software at Nvidia, explained that “this is really agentic behavior that we’re talking about, which is fleets of agents and how they operate together.”
See Also: North Korean Hackers Cross $1 Billion Mark After Huge Crypto Theft
Nvidia leadership emphasized that these protective measures could prevent major cyberattacks before they happen. Justin Boitano, vice president and general manager of enterprise computing at Nvidia, pointed directly to recent real world digital breaches during a media briefing.
He noted that “from what we know, this new security platform could have stopped the breach if it was being used in frontier labs for model evaluation early on,” adding that “we’re advancing this openly, and we want to engage everybody to work with us.”
Rather than keeping the code secret, Nvidia made the entire security framework open-source and free for anyone to download and modify.
Major software developers, robotics creators, and AI laboratories have already started integrating the guardrails into their products.
By offering these free tools, Nvidia hopes to make safety a standard across the entire technology industry, ensuring that as artificial intelligence becomes more powerful, human owners remain in full control.

