Nvidia says its new AI safety platform can contain rogue agents within ‘milliseconds’

The Open Agent Safety Platform is designed to enforce AI boundaries.


Nvidia is launching a new safety platform designed to contain and monitor AI agents, a move that comes in response to a wave of rogue hacking incidents , as reported earlier by Reuters . In an announcement on Monday , Nvidia says its new Open Agent Safety Platform can quarantine agents that attempt to escape their boundaries within “milliseconds.”
The platform uses Nvidia’s OpenShell open-source software, which runs on the company’s Vera AI CPU . Users can choose the information an AI agent can access, and OpenShell checks these restrictions before and during a task, according to Nvidia . It also includes Nvidia’s Sentry technology on a separate chip to continuously monitor agents and enforce boundaries.
During an interview with CNBC , Nvidia CEO Jensen Huang emphasized the importance of giving AI agents access to only the information they need to do their job. “In order for you to deliver that agentic system in a safe way, you have to make sure that the sandbox around it… all of those systems are designed in a way that keeps the agent with minimal rights,” Huang said.
Concerns about AI safety have risen in recent weeks , as OpenAI , Anthropic , and Google have all revealed incidents where their AI models went outside their testing environments and hacked other companies. Several major tech companies are backing Nvidia’s Open Agent Safety Platform, including Anthropic, Microsoft, and SpaceX.
Update, September 28th: Added Huang’s CNBC interview.
More in: The AI Superintelligence Slowdown
Verified source · The Verge
Reported by The Verge. Open the original for full media and formatting.
More in Policy
All news
PolicyOpenAI reportedly ditches model over safety concerns
A top executive at the AI lab told the Wall Street Journal that the model in question had displayed a poor aptitude for following orders.
Read at TechCrunch
PolicyTrump finalizes rule to make cars less fuel efficient
The US Department of Transportation finalized its plans today to weaken fuel efficiency standards, calling it "among the largest deregulatory actions under the second Trump Administration." It's a nail in the coffin for Biden-era standards that would have required fleet average…
Read at The Verge
PolicyAI is supercharging hacking, and your local hospitals and banks aren’t ready
In March, Janice Malone began getting calls about suspicious activity from her nonprofit organization, Vivian's Door. Vivian's Door, headquartered in Alabama, typically provided training, resources, and community to underserved and minority-owned businesses. The work sometimes p…
Read at The Verge
PolicyOpenAI still doesn’t seem to have a handle on all of its rogue AI activity
On Friday, OpenAI published a new site devoted to “misalignment reports” and the breadth of the incidents is alarming.
Read at TechCrunch