Nvidia has launched an Open Agent Safety Platform designed to constrain autonomous systems even when application-level safeguards fail. The company says the platform combines OpenShell, an open-source runtime that traces actions and enforces policy, with Sentry, a reference design that uses BlueField-4 data-processing units as an independent watchdog.
The architectural point matters more than the product branding. OpenShell is intended to isolate agents and control their access to models, tools and data. Sentry sits outside the agent and host software, correlating activity and quarantining agents that cross defined boundaries. Nvidia says OpenShell can be extended to Arm and Intel environments, while Sentry is tied to its BlueField hardware. Those claims will need independent testing, particularly around policy bypasses, latency and operational failure modes.
The launch also creates an unusually broad enterprise coalition. Nvidia named Anthropic, Cisco, CrowdStrike, HPE, JPMorganChase, Microsoft, Palantir, Palo Alto Networks, Salesforce and ServiceNow among participating organisations. For regulated firms, the important question is whether out-of-band enforcement can provide stronger evidence of control than prompt-layer guardrails alone. The platform is therefore best read as a proposed control plane for agentic systems, not as proof that autonomous deployment is now safe.
