Nvidia launched a tool designed to stop AI agents from going rogue. Here’s how it works.
CEO Jensen Huang said keeping AI agents in line is an engineering problem Nvidia can solve. Benjamin Fanjoy/Getty Images AI agents have gone out of whack, escaping tests, accessing systems, and covering it up. Nvidia wants to stop it by fencing agents in and quickly cutting them off when they stray. CEO Jensen Huang framed it as a "technically solvable problem" of engineering. Nvidia is giving AI agents a playpen — and a watchdog. The chipmaker on Monday launched the Open Agent Safety Platform, a two-part system designed to keep agents inside clearly defined boundaries and cut them off quickly if they try to escape. It comes after several frontier AI labs reported agents breaking out of supposedly secure testing environments, reaching systems they were not meant to access, and sometimes misrepresenting what they had done.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in