Start your day with intelligence. Get The OODA Daily Pulse.

Home > Briefs > Technology > Anthropic says its Claude models ‘gained unauthorized access’ to other organizations’ systems

Anthropic says its Claude models ‘gained unauthorized access’ to other organizations’ systems

Anthropic on Thursday said it discovered three instances where its Claude artificial intelligence models accessed the internet during an evaluation and “gained unauthorized access to the real systems of three different organizations.” The company said it found these incidents after carrying out a “a large-scale retrospective review” of its cybersecurity evaluations. Anthropic said the review was prompted by a separate but similar security incident that OpenAI disclosed last week. OpenAI said a combination of its models escaped an isolated testing environment that had very limited internet access. The models chained together a series of vulnerabilities to reach the open web and eventually gain access to Hugging Face, which operates an open-source developer platform. In the three incidents that Anthropic detected, its models accessed the internet while interacting with a testing environment from one of its third-party evaluation partners called Irregular.

Full report : Anthropic Says Its A.I. Systems Broke Into Computers at 3 Organizations.