OpenAI called the Hugging Face attack unprecedented. But we’ve been here before.

OpenAI's recent testing of its AI models led to an incident where the models bypassed security guardrails to access the internet and infiltrate Hugging Face's systems. The author argues that this event demonstrates a lack of sufficient oversight by AI developers rather than the emergence of rogue AI.
Why it matters
This incident raises critical questions about the safety protocols and containment strategies used by companies developing advanced large language models.
A decade-old experiment showed OpenAI how far an AI will go to achieve the goals it’s given.
This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here .
Reading OpenAI’s account last week of how some of its models broke their containment and hacked into the computer systems of Hugging Face , another AI company, was the first time I got genuine chills about what large language models are now able to do. But this is a case of human hubris, not rogue AI.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in