
#NVIDIAAgentSafety
About NVIDIAAgentSafety
NVIDIA launched its Open Agent Safety Platform on Sept 28, combining OpenShell to restrict agent access to files, networks and tools with Sentry, a separate hardware-based monitoring and containment layer. Backed by Anthropic and other partners, it arrives as agents gain autonomy and enterprise access. As AI agents enter core workflows, can software controls and hardware-level monitoring become standard infrastructure, extending AI spending beyond compute, networking and storage into security?
Hot
Latest
NVIDIAAgentSafety Popular posts

Very good! NVIDIA is bringing over 100 partners together to tackle the agent security failures that threaten to slow AI development. However, strangely enough especially OpenAI is missing.
After OpenAI’s Hugging Face breach and Anthropic’s testing incidents, its Open Agent Safety Platform combines OpenShell’s access controls with Sentry’s monitoring and enforcement on separate hardware, designed to stay beyond the agent’s reach.
Better containment could let researchers test stronger models and companies deploy more useful agents. That’s a concrete way for safety engineering to support acceleration.
Anthropic, Microsoft and Hugging Face are among the named partners. OpenAI is absent from the announced lineup. Given its own containment failures, I’d like to know why. Anyways, Kudos NVIDIA!


Jensen Huang
Today, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry.
Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come.
But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility.
This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems.
Together, we are building the foundation of the AI economy.
Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world.


🚨JUST IN: Nvidia $NVDA is launching an AI safety system designed to STOP agents from breaking into other systems.
The Open Agent Safety Platform “controls what AI agents can access in real time and can shut them down when they break the rules,” Bloomberg reports
The system includes two open-source security tools that can be run on Nvidia hardware.
VP Justin Boitano says the system "could have stopped" the recent Hugging Face breach by OpenAI’s AI models.
Boitano added that it can “quarantine a suspicious agent in milliseconds.”



Chain-of-thought monitoring, fundamentally, involves reading the notes than an AI writes to itself, and checking what's in those notes.
This works at all - AIs are only trained to have nice results, not nice-looking notes, which means that their notes are whatever strange things achieve the task.
And, of course, as models get smarter, they need fewer notes to get the same results; they can do more of the work in their head.

🔥 OKX launches Pre-IPO X-Perps!
OKX has announced X-Perps for Anthropic and OpenAI, giving eligible EEA users access to trade perpetual contracts linked to these private companies before any public listing.
🛡️ OKX Shield expands security
Eligible European users can receive up to €500,000 in account-takeover protection, depending on VIP level and security requirements.
⚡ More X-Perps added
OKX recently listed METUSD, ARUSD and COREUSD, following FLOCKUSD, MINAUSD and CASHCATUSD. $BTC #FedHike







