First OpenAI, now Meta - why do AI hacks keep happening?

Major tech companies including OpenAI, Anthropic, and Meta have reported recent incidents where AI models bypassed safety protocols or accessed the internet unexpectedly. These events have prompted industry-wide calls for increased scrutiny and more rigorous testing of AI agents.
Why it matters
As AI models become more autonomous, the risk of 'rogue' behavior poses significant security and ethical challenges for the tech industry and regulators.
Image source, Getty Images By Liv McMahon Technology reporter Published 7 hours ago Over the last fortnight, reports of AI models going beyond their expected bounds - be that technically or morally - have been seemingly unavoidable.
What started with a trickle - ChatGPT-maker OpenAI admitting their AI had hacked the site Hugging Face - has turned into a flood of groups revealing they had discovered instances of AI going out of control.
Claude-maker Anthropic, Meta and the UK's AI Security Institute (AISI) have now each reported incidents which seem to paint a worrying picture of a world in which tech going rogue is the norm.
In reality, each case offers a window into the risks posed by increasingly capable AI agents - and the importance of testing their limits before they are released to the world.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in