Meta joins OpenAI, Anthropic: AI model went 'wild', blames misconfiguration
Meta has disclosed that its Muse Spark AI model accessed the internet during testing due to a misconfiguration by a third-party firm. This incident follows similar reports from OpenAI and Anthropic, sparking broader industry discussions regarding AI safety and transparency.
Meta has confirmed that its Muse Spark model exploited a security vulnerability in a third-party service during cybersecurity testing, marking the company’s first public disclosure of a rouge AI incident. According to a report by Business Insider, in a statement, Meta said that the breach occurred due to a misconfiguration by Irregular, an independent firm it uses for model evaluations. The error allowed the model to access the internet during testing. Meta added that it learned of the incident when Irregular notified the company and is now investigating, promising a full retrospective once details are finalised.Meta becomes the third AI company to report such incidentsMeta’s disclosure makes it the third major AI company to report such incidents in recent weeks. OpenAI admitted that two of its models escaped test environments and hacked into Hugging Face, later self-reporting two additional lapses. Anthropic revealed that its Claude models had gained unauthorised access to external systems during testing. Hugging Face itself confirmed a breach last month, while other firms have faced similar rogue agent behaviour.What Irregular said about Meta’s Muse Spark breachIrregular, the testing company involved, said the incident was not a sophisticated cyberattack but rather an evaluation environment issue similar to Anthropic’s disclosure. “This did not involve a sandbox escape or a sophisticated cyber action. There are no current open issues,” a spokesperson said, adding that Irregular is preparing a white paper on best practices for containment and secure evaluations.Calls for transparencyThe string of incidents has prompted calls for stronger AI safety regulation and mandatory disclosures. Hugging Face CEO Clem Delangue told CBS that transparency is critical, urging companies to share “agent traces” to determine whether lapses stemmed from human error, system flaws, or AI behavior. Industry leaders like Box CEO Aaron Levie have warned these breaches highlight the “wild times” ahead, as AI agents demonstrate the ability to escape systems, discover vulnerabilities, and hack external platforms.Get the latest technology news and updates. Download the TOI App.
Get the full story
Sign up for Headlinne to unlock AI insights, political bias analysis, and your personalized news feed.
Create free accountAlready have an account? Sign in