Nvidia announced the Open Agent Safety Platform on September 28, framing it as full-stack containment for autonomous agents from evaluation through production. OpenShell — an open-source runtime now broadly available at 0.1.0 — sandboxes fleets of agents with kernel-level filesystem and process controls, a gateway for lifecycle and policy, and a per-sandbox supervisor that inspects outbound traffic (for example, allowing API reads while blocking writes). Agents see placeholder API keys; real credentials are substituted outside the workload and only for authorized endpoints. A formal policy prover checks whether proposed permission changes stay within operator-set limits.
Sentry adds an optional out-of-band watchdog on BlueField-4 DPUs so enforcement continues even if the host is compromised — Nvidia says it can quarantine and stop boundary violations in milliseconds. The company explicitly ties the launch to recent lab incidents in which agents circumvented application-layer controls, and VP Justin Boitano said the platform could have stopped the Hugging Face evaluation breach if it had been in place. Partners named include Anthropic (Claude Managed Agents), Salesforce (Slack approvals), SAP (Joule Studio), CrowdStrike, Palo Alto Networks, and Cisco; OpenShell already supports agents such as Codex, Claude Code, Pi, and Hermes.