Nvidia has recently introduced a new security platform known as Open Agent Safety Platform to control AI agents as they become capable of taking more actions on their own. The company claims that the new tool can keep systems within defined limits and prevent them from accessing tools, files or computer systems without permission. Unlike chatbots, AI agents can browse the internet, use software and complete tasks over longer periods, creating new security concerns. Nvidia says its platform combines software and hardware controls to watch agent activity and stop risky actions. Here’s everything you need to know about all new Nvidia Open Agent Safety Platform.
Nvidia CEO Jensen Huang recently shared a post on its X handle introducing the Nvidia Open Agent Safety Platform. According to the details, the platform has two main parts, notably the OpenShell and Sentry. OpenShell is open-source software that creates a controlled environment for an AI agent and lets developers set rules around what an agent can access. Not only that, but the company also clarified that OpenShell also works with platforms based on ARM and Intel chips.
Today, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry.
— Jensen Huang (@JensenHuang) September 28, 2026
Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to… pic.twitter.com/dReAxwpRUn
Whereas the Sentry adds another layer of protection at the hardware level. It runs on Nvidia’s BlueField-4 data processing units and separately watches what an AI agent is doing. If an agent tries to move outside its permitted environment, Sentry can isolate and stop it within milliseconds, according to Nvidia.
Also read: Apple iPhone 18 price in India, launch timeline, camera, display and all other leaks
The post also confirms that the company is working with more than 100 organisations on the platform, which include the big names like Anthropic, Microsoft, Cisco, CrowdStrike, Dell, HPE, Hugging Face, Palantir, Salesforce, SAP, Scale AI and ServiceNow.
While announcing the platform in the post, Huang said, ‘Trust and innovation are not in conflict. Safety is how trust is earned.’ Not only that, but while pointing to recent incidents involving OpenAI where the company left its testing environment accessible to the open internet, allowing the system to interact with infrastructure linked to Hugging Face, Nvidia says that their new technology could have helped prevent such incidents.
Also read: Redmi 17C India launch this week: Check expected specs, price and more
Justin Boitano, Nvidia’s vice president of enterprise AI, said Hugging Face reported more than 17,000 agents attacking its infrastructure over the course of days and weeks. He further added that model-level safety measures alone cannot fully control what an AI agent can access once it is connected to external tools and systems. Nvidia is presenting the platform as an open reference design.
Anthropic is working with Nvidia to add OpenShell and BlueField-based controls to its managed AI agents. The software is available through Nvidia’s developer resources and GitHub as AI agents become more widely used.