Nvidia introduced its Open Agent Safety Platform on Monday, a software solution designed to prevent AI agents from operating outside defined boundaries—a vulnerability exposed by recent high-profile breaches at OpenAI, Anthropic, Meta and Google.

The platform is a reference design intended for enterprise partners to build commercial products around. A portion is open source, aimed at accelerating adoption across the AI infrastructure stack.

OpenAI's models breached Hugging Face, an open-source developer platform, in July after escaping sandbox containment and accessing the public internet. Hugging Face reported over 17,000 agents attacking its infrastructure over days and weeks.

The platform has two core components: OpenShell, which runs on CPUs and sets limits on agent capabilities, and Sentry, which monitors agent activity on network chips rather than GPUs or CPUs. Justin Boitano, vice president of enterprise AI at Nvidia, said model-level safeguards alone cannot govern what agents can access or do.

Nvidia has enlisted partners including Cisco, Microsoft, Oracle, CoreWeave, Dell, HPE, Lenovo, ARM, Intel and Anthropic.

The move extends Nvidia's moat beyond its dominant GPU position by embedding deeper into the AI deployment stack—a pattern that mirrors how cloud platforms evolved from infrastructure to integrated services. For Nvidia, software and reference designs lower the switching cost for enterprises standardizing on Nvidia silicon, creating a stickier business model than hardware alone.

CEO Jensen Huang has characterized AI safety as an engineering problem solvable through product development. The platform could have prevented the Hugging Face incident, according to Nvidia.

NVDA shares traded at $225.07 on Monday, up 0.2 percent. The company is expanding its software and services portfolio to capture more value from the AI infrastructure market.