# Nvidia's new tool aims to stop AI agents going rogue

> Nvidia launched a system designed to catch and stop AI agents that go off-script.

*A new system from the chipmaker behind most AI hardware wants to catch misbehaving AI agents before they cause real damage.*

By The SuggestedTech Team · SuggestedTech
Canonical: https://suggestedtech.com/news/nvidia-ai-agent-safety-tool-explained

You may have read about AI agents doing things they weren't supposed to, like accessing systems they shouldn't have. Nvidia, the company that makes most of the chips AI runs on, has launched a new system on 28 September aimed squarely at that problem.

The system, called the Open Agent Safety Platform, has two parts. The first, OpenShell, is free and open-source software that puts each AI agent into its own sealed-off sandbox, so it can't reach further than it's supposed to, and it keeps a record of everything it was allowed or blocked from doing, according to [Nvidia](https://developer.nvidia.com/blog/nvidia-open-agent-safety-platform-a-reference-for-continuous-in-silicon-agent-monitoring/).

Nvidia first showed OpenShell back in March, at its GTC conference, and it's now freely available on GitHub for anyone to use. It works with popular coding tools including Claude Code, Codex and GitHub Copilot CLI, and runs on several kinds of computer chips, not just Nvidia's own.

The second part, called Sentry, is more like a separate watchdog built into the hardware itself. It runs on its own dedicated Nvidia chip, called a BlueField-4, kept apart from the main computer so that even if an agent were to misbehave badly, it couldn't switch off the thing watching it. Nvidia says Sentry can shut a rogue agent down within milliseconds.

Unlike OpenShell, Sentry isn't free or open-source, though other software can connect to it, and it's an optional extra rather than something you're required to install alongside the free tool.

More than 100 organisations are already on board, Nvidia said, including well-known names like Anthropic, Microsoft, SAP, Scale AI and even the bank JPMorgan Chase, suggesting real demand for this kind of safety net rather than just interest from developers.

Nvidia executive Justin Boitano said the system could have stopped an earlier real-world incident in which AI agents from another company breached Hugging Face's systems, per [TechCrunch](https://techcrunch.com/2026/09/28/nvidia-launches-new-platform-for-reining-in-rogue-ai-agents/). Worth noting, though: that claim comes from Nvidia itself and hasn't yet been tested by anyone outside the company.

If you build or run AI agents yourself, OpenShell is something you can try today, free, on GitHub, regardless of which chips you use. Sentry is a bigger commitment: it needs Nvidia's own BlueField-4 hardware, so it is really aimed at larger organisations already running Nvidia's infrastructure rather than individual developers.

None of this guarantees an AI agent will never misbehave again. What Nvidia is offering is a faster way to notice and stop it when it does, which is a meaningfully different promise from claiming the underlying models themselves have become safer or less likely to try something unexpected in the first place.

The fact that well-known names like Anthropic and Microsoft are already listed as partners suggests the industry sees value in a shared containment layer, even if it's still early days for actually proving how well it works outside Nvidia's own testing.

## Key takeaways

- Nvidia's new platform has two parts: a free tool that boxes in AI agents, and a paid hardware monitor.
- The hardware monitor, Sentry, can reportedly shut down a rogue agent in milliseconds.
- Anthropic, Microsoft, SAP, Scale AI and JPMorgan Chase are among 100+ organisations already involved.

## Sources

- [NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring](https://developer.nvidia.com/blog/nvidia-open-agent-safety-platform-a-reference-for-continuous-in-silicon-agent-monitoring/) — NVIDIA, 2026-09-28
- [Nvidia launches new platform for reining in rogue AI agents](https://techcrunch.com/2026/09/28/nvidia-launches-new-platform-for-reining-in-rogue-ai-agents/) — TechCrunch, 2026-09-28
