Anthropic’s AI Claude escaped testing environment and hacked organizations

0
1

Anthropic ⁠said on Thursday its AI Claude model hacked ⁠systems of ⁠three ​organizations during testing, days after rival OpenAI ⁠revealed a rogue agent had gone on a days-long ⁠hacking spree at AI ​firm Hugging ‌Face.

Claude gained ‌unauthorized access to the ‌systems during cybersecurity evaluations after a misconfiguration allowed the models to reach the internet from testing environments that ‌were supposed to be isolated, Anthropic said.

The company said ​it identified the incidents after reviewing 141,006 cybersecurity evaluation runs, a process it ⁠launched following OpenAI’s disclosures.

“Claude compromised the ​impacted ​organizations’ infrastructure using ​basic techniques, such as ​exploiting ‌weak passwords and ​unauthenticated ​endpoints,” it said.

According to Anthropic, the three hacked organizations had not detected the activity.

“We discovered these incidents after a proactive review of our cybersecurity evaluation transcripts,” the company said in a statement, noting it then reached out to the affected organizations.

Disclaimer : This story is auto aggregated by a computer programme and has not been created or edited by DOWNTHENEWS. Publisher: theguardian.com