Technology

Anthropic admits that its AI breached the systems of three companies during a test

It happened during the testing phase. The incident came just a few days after OpenAI’s warning

FILE PHOTO: FILE PHOTO: Anthropic logo is seen in this illustration taken May 20, 2024. REUTERS/Dado Ruvic/Illustration/File Photo/File Photo REUTERS

2' min read

Translated by AI
Versione italiana

2' min read

Translated by AI
Versione italiana

Anthropic has stated that its artificial intelligence models breached the systems of three other organisations during testing, just days after OpenAI, the company behind ChatGPT, had raised concerns about AI control following the revelation that its unauthorised models had breached another company’s systems.

The case

Anthropic, the San Francisco-based AI company that created Claude, announced on its website on Thursday that it had discovered the three incidents after examining over 141,000 evaluation runs.

Loading...

In response to the OpenAI incident, Anthropic stated that it had launched a “large-scale” cybersecurity review specifically aimed at verifying whether its AI models had been able to access the internet within test environments that were supposed to be isolated.

Anthropic has clarified that the models involved in the incidents are Claude Opus 4.7, Claude Mythos 5 and an internal test model used for research. The first incidents date back to April, the AI company said.

“Claude compromised the infrastructure of the affected organisations using basic techniques,” said Anthropic, such as exploiting weak passwords.

Countermeasures

The company added that it had already contacted the organisations involved – the names of which it did not disclose – and that two of them had stated they had not previously detected the activity, whilst the artificial intelligence company “continues to contact the third”.

Last week, OpenAI stated that its artificial intelligence models had run amok during a testing phase, breaching the servers of the AI start-up Hugging Face. OpenAI described the incident as a “serious security incident”.

These incidents have highlighted vulnerabilities in AI security and control, and raised questions about how AI can be managed safely by humans as its use becomes more widespread globally.

“Safety tests are carried out before a model is released precisely because we do not yet know what it is capable of,” Anthropic stated on its website on Thursday.

Copyright reserved ©
Loading...

Brand connect

Loading...

Newsletter

Notizie e approfondimenti sugli avvenimenti politici, economici e finanziari.

Iscriviti