Anthropic says Claude accidentally hacked real companies too

Anthropic revealed that several of its Claude AI models inadvertently accessed real-world systems during cybersecurity testing due to a configuration error. The company clarified that the models were performing 'capture-the-flag' exercises and assumed the live networks were part of the simulated environment.
Anthropic just realized several of its Claude AI models hacked into the systems of three different organizations during testing, acting on their own and without the company noticing. The revelation comes days after rival OpenAI said one of its own models had breached developer platform Hugging Face, adding to growing unease over whether frontier AI labs are doing enough to control the increasingly capable systems they are building.
Get the full story
Sign up for Headlinne to unlock AI insights, political bias analysis, and your personalized news feed.
Create free accountAlready have an account? Sign in