Anthropic says its AI models hacked systems of three companies during tests

Anthropic's AI model Claude hacked systems of three organizations during testing due to a misconfiguration that allowed it to access the internet from isolated testing environments.

Anthropic says its AI models hacked systems of three companies during tests
This image is AI-generated and does not depict any real-life event or location. It is a fictional representation created for illustrative purposes only.
  • Country:
  • United States

Anthropic ​said on ‌Thursday its ​AI Claude model hacked systems of three ‌organizations during testing, days after rival OpenAI revealed a rogue agent had gone on ‌a days-long hacking spree at AI ‌firm Hugging Face.

Claude gained unauthorized access to the systems during cybersecurity evaluations after a ⁠misconfiguration ​allowed the models ⁠to reach the internet from testing environments ⁠that were supposed to be isolated, Anthropic ​said. The company said it identified the ⁠incidents after reviewing 141,006 cybersecurity evaluation runs, ⁠a ​process it launched following OpenAI's disclosures.

"Claude compromised the impacted organizations' infrastructure using ⁠basic techniques, such as exploiting weak passwords and ⁠unauthenticated ⁠endpoints," it said.

Give Feedback

Use this form for editorial or site feedback. We usually reply within 2 to 3 working days.

By submitting, you agree that we may use your email address to respond.