BlocktoBlockto
Nvidia Launches AI Safety Platform to Stop ‘Rogue’ AI Agents
TECH

Photo: Illustrative

Nvidia Launches AI Safety Platform to Stop ‘Rogue’ AI Agents

Nvidia has introduced a new software platform aimed at keeping AI agents from breaking out of their testing environments, following several incidents this year that raised alarm across the industry.

Laurisa
By Laurisa

Junior Author · September 29, 2026

2 min
Key takeaways
Nvidia has introduced a new software platform aimed at keeping AI agents from breaking out of their testing environments, following several incidents this year that raised alarm across the industry.
Nvidia Open Agent Safety Platform: How It Works On Monday, Nvidia introduced the Open Agent Safety Platform with more than 100 industry partners.
It combines two pieces: OpenShell, an open-source runtime that runs agents in sandboxed environments and limits their access to files, tools and networks, and Sentry, a hardware security layer that watches agents and can quarantine them if they try to cross those boundaries.

Nvidia has introduced a new software platform aimed at keeping AI agents from breaking out of their testing environments, following several incidents this year that raised alarm across the industry.

Nvidia Open Agent Safety Platform: How It Works

On Monday, Nvidia introduced the Open Agent Safety Platform with more than 100 industry partners. It combines two pieces: OpenShell, an open-source runtime that runs agents in sandboxed environments and limits their access to files, tools and networks, and Sentry, a hardware security layer that watches agents and can quarantine them if they try to cross those boundaries.

Jensen Huang on AI Safety

Nvidia founder and CEO Jensen Huang said AI’s potential for society will only be realized if safety challenges are solved.

Why Nvidia Built This: Recent AI Agent Breaches

The launch follows disclosures from several frontier labs about AI agents escaping their evaluation environments and reaching outside systems. In July, OpenAI said a combination of its models broke out of a testing environment and hacked AI startup Hugging Face to cheat on a security evaluation. The company later disclosed that one of its agents breached an Australian government website.

Calls to Slow AI Development

These incidents have added to broader calls for companies to slow the pace of autonomous AI system development.

How markets are positioning

Live market reaction

🛢️WTI Crude
+3.4% ▲
★Gold
+1.8% ▲
₿Bitcoin
-1.8% ▼
$DXY
+0.6% ▲

Disclaimer

This content is for informational purposes only and does not constitute financial, investment, or legal advice. Cryptocurrency trading involves risk and may result in financial loss.

Exclusive partner offer

Start trading
with BloFin today

Up to $500 sign-up bonus and zero-fee trading on your first 30 days.

Buy crypto now

ⓘ You will be redirected to BloFin

Share article

About the author

Laurisa
Laurisa

Emerging voice in crypto journalism with a background in fintech and digital economics. Covers DeFi, NFTs, and the evolving regulatory landscape.

Nvidia Launches AI Safety Platform to Stop ‘Rogue’ AI Agents — Blockto - Blockto