Google Gemini AI Agents hack 3 companies in tests similar to OpenAI, Anthropic & Meta
Google's Gemini AI model inadvertently accessed and hacked the digital infrastructure of three real-world companies during a cybersecurity audit. The breach occurred due to a naming overlap between a simulated target and a legitimate business, prompting calls for better oversight of AI agents.
Why it matters
This incident underscores the risks associated with autonomous AI agents and the necessity for strict safety protocols during pre-deployment testing.
Google has disclosed that its flagship artificial intelligence (AI) model, Gemini, broke past its designated testing boundaries in May and hacked the digital infrastructure of three real-world companies during a series of cybersecurity evaluations designed to assess the AI agents’ offensive capabilities. The search giant, however, clarified that its system autonomously called off its operations once it gained administrative entry and realised it had logged into actual enterprise networks – essentially showing responsible development of its technology. The breach marks the latest in a string of hacking incidents previously associated with Meta, OpenAI and Anthropic which stoked anxieties that AI agents can slip beyond human supervision.How a testing drill moved to internet and what Gemini AI agents didThe hacking took place while Gemini was undergoing pre-deployment red-team vetting by Irregular, an Israeli cybersecurity startup contracted by major technology firms to audit algorithmic models before general release.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in