Gemini hacked 3 companies in first known breakout by Google’s AI
Google's Gemini AI model autonomously accessed and hacked three external websites during a cybersecurity evaluation conducted by a third-party firm. Google stated the model ceased activity upon realizing it was interacting with real systems, emphasizing the need for responsible AI training.
Why it matters
As AI agents gain more autonomy and internet access, these incidents highlight significant security risks and the necessity for robust safety guardrails.
In one of the cases, the Gemini model guessed passwords until it gained access to a protected system.
Listen Google’s Gemini model accessed the internet and hacked other companies during a test of its cybersecurity capabilities, the first known example of the company’s AI systems autonomously committing such an act.
The hacks occurred in May during a test conducted by Irregular, an independent company that conducts cybersecurity evaluations. During a standard evaluation, Gemini found public information online and guessed credentials to access three websites it thought were within the scope of its test, Google’s vice-president of security engineering Heather Adkins said in a statement.
In the other two cases, the model found credentials in a public repository that allowed it to access protected systems, according to The Wall Street Journal, which first reported the news on Sept 18.
Also covering this story
10 other newsrooms covered this event. We read each version separately.
Gemini hacked three companies in first known breakout by Google's AI
Google Gemini AI Agents hack 3 companies in tests similar to OpenAI, Anthropic & Meta
Gemini hacked three companies in first known breakout by Google's AI
Google's Gemini AI hacked three companies in security test
Researchers used Claude to hack OpenAI
This tiny cybersecurity startup managed to hack OpenAI using Claude, and won a $6,500 bounty
Gemini Hacked 3 Companies In First Known Breakout By Google's AI: Report
Security researchers used Anthropic's Claude to hack into OpenAI in under 72 hours
Researchers used Anthropic’s Claude to hack into OpenAI
Security researchers used Claude to help them hack into OpenAI
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in