After OpenAI agent 'caught hacking', chief scientist warns: Firms not prepared
OpenAI chief scientist Jakub Pachocki has warned that AI agents are increasingly capable of hacking and deceptive behaviors. He emphasized that current safeguards are insufficient to prevent autonomous systems from evading oversight.
Why it matters
As AI agents become more autonomous, the potential for them to manipulate systems or humans poses a significant risk to global digital infrastructure.
OpenAI recently introduced its newest model Astra. Now days after the launch ChatGPT-maker has once again discovered its AI agents engaging in hacking behaviour. The incident prompted chief scientists Jakub Pachocki to issue a stark warning: “No one is prepared for the consequences of a continued rapid rise in machine intelligence.” Pachocki said that while OpenAI is working on technical safeguards, broader interventions are needed to prevent autonomous agents from evading oversight, breaking into systems, or tricking people to achieve their objectives.The risks OpenAI chief scientist Jakub Pachocki is warning aboutAgents that can hack, deceive, and manipulatePachocki said AI agents are becoming exceptionally skilled at breaking into protected systems across the open internet, putting global infrastructure at risk. He argued that there's currently only a narrow window to use today's best models to substantially strengthen the security of critical systems before that risk grows further.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in