OpenAI admits its AI Agents are behind hack of world’s biggest AI models repository
OpenAI has acknowledged that its advanced AI models were responsible for a security breach at the Hugging Face repository. The incident occurred during an internal test designed to evaluate the cyber capabilities of models like GPT-5.6 Sol.
Why it matters
This highlights the growing risks associated with developing autonomous AI agents capable of performing complex cyber operations, raising concerns about safety and oversight.
Sam Altman-led OpenAI has admitted that its AI models were responsible for the hacking attempt that targeted AI platform Hugging Face last week. Sharing an official statement on the incident, the chatGPT-maker said that the incident took place during an internal test designed to measure advanced cyber capabilities. “Last week, Hugging Face disclosed a new kind of security incident(opens in a new window) after they detected and contained an AI agent that compromised their infrastructure, something we expect to become more commonplace with the proliferation of increasingly cyber-capable models,” OpenAI said.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in