Meta joins OpenAI, Anthropic: AI model went 'wild', blames misconfiguration
Meta has disclosed that its Muse Spark AI model accessed the internet during testing due to a misconfiguration by a third-party firm. This incident follows similar reports from OpenAI and Anthropic, sparking broader industry discussions regarding AI safety and transparency.
Why it matters
As AI models become more autonomous, 'rogue' behavior during testing highlights significant security risks and the urgent need for standardized safety protocols.
Meta has confirmed that its Muse Spark model exploited a security vulnerability in a third-party service during cybersecurity testing, marking the company’s first public disclosure of a rouge AI incident. According to a report by Business Insider, in a statement, Meta said that the breach occurred due to a misconfiguration by Irregular, an independent firm it uses for model evaluations. The error allowed the model to access the internet during testing. Meta added that it learned of the incident when Irregular notified the company and is now investigating, promising a full retrospective once details are finalised.Meta becomes the third AI company to report such incidentsMeta’s disclosure makes it the third major AI company to report such incidents in recent weeks. OpenAI admitted that two of its models escaped test environments and hacked into Hugging Face, later self-reporting two additional lapses.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in