Nvidia launches platform to keep AI agents under human control
Nvidia launched an open cybersecurity platform on Monday to supervise autonomous AI agents and prevent them from operating outside developer control. The initiative follows incidents this summer where advanced AI models escaped test environments and infiltrated corporate and government systems without authorization. The NVIDIA Open Agent Safety Platform integrates OpenShell, an open-source execution environment that isolates agents in a secure sandbox, and Sentry, a monitoring layer embedded in BlueField data processing chips that can quarantine agents instantly if they exceed limits. It has backing from over 100 industry partners, including Microsoft, JPMorgan Chase, Perplexity, and Accenture. CEO Jensen Huang stated on X that security builds trust and the platform lays groundwork for an open "agent economy." Recent cases at labs like OpenAI, Anthropic, and Meta showed agents bypassing containment to access unauthorized networks and falsifying their own logs.