A A A
Temperature Humidity
News Archive Can search within past 12 months

Anthropic models hacked three systems during tests

2026-07-31 HKT 08:07
Share this story facebook
  • Anthropic said a review prompted by revelations of a rogue agent operated by rival OpenAI found three incidents where its own AI agent gained "unauthorised access" to the systems of three different organisations. Image: Reuters
    Anthropic said a review prompted by revelations of a rogue agent operated by rival OpenAI found three incidents where its own AI agent gained "unauthorised access" to the systems of three different organisations. Image: Reuters
Anthropic said on Thursday its AI Claude model hacked ⁠systems of three organisations during testing, days after rival OpenAI revealed a rogue agent had gone on a days-long hacking spree at AI ⁠firm Hugging Face.

Claude ⁠gained unauthorised access ⁠to the systems during cybersecurity evaluations after ⁠a misconfiguration allowed the models to reach the internet from testing environments that were supposed to ⁠be isolated, Anthropic said.

The company said it identified the incidents after reviewing 141,006 cybersecurity evaluation runs, a process it launched following OpenAI's disclosures.

"Claude ⁠compromised the impacted organisations' infrastructure using basic techniques, such as exploiting weak passwords and unauthenticated endpoints," it said. (Reuters)



Edited by Cecil Wong

Anthropic models hacked three systems during tests