Nvidia rolls out Open Agent Safety Platform, moving AI agent controls to the chip layer
Nvidia has introduced the Open Agent Safety Platform, a new security stack built to place guardrails around AI agents at both the software and hardware levels. The platform has two components: OpenShell, which defines and enforces permission boundaries for agents, and Sentry, a newly added hardware security layer that runs on the BlueField-4 chip side. Nvidia said an agent that steps outside its allowed boundaries can be isolated within milliseconds. OpenShell had already been open-sourced earlier and can be wrapped around existing agents such as Claude Code and Codex without changing the agents themselves. It is designed to restrict file access, network connections, tool usage, and credential use, while enforcing those policies during runtime. Sentry operates independently from the agent, continuously watching behavior, tool calls, and data access, and handling threat detection and policy enforcement even if the agent behaves abnormally. Jensen Huang said more than 100 industry partners are already involved in the platform, including Anthropic, Microsoft, Salesforce, SAP, SpaceX, CrowdStrike, and Palantir. Nvidia framed the launch as a response to agents running for hours or even days at a time while interacting with code, tools, and enterprise data.

