The announcement

Anthropic: AI systems have carried out ‘unintended’ actions on government websites

During internal tests, Anthropic’s AI models carried out unauthorised operations on federal and state portals, prompting the company to suspend their internet access

2' min read

Translated by AI
Versione italiana

2' min read

Translated by AI
Versione italiana

Anthropic, the company behind the Claude chatbot, has stated that some of its AI agents carried out unintended actions on federal, state and local government websites, including sending a false report of a murder to the Philadelphia police emergency hotline.

The internal investigation

This is according to the Washington Post. The US State Department announced yesterday that a test model had submitted 19 visa applications in August and another in May using a form on the department’s website. Anthropic stated that an internal review had revealed that its artificial intelligence models had behaved inappropriately in some instances during the testing phase and whilst being used by employees.

Loading...
Difesa e Intelligenza artificiale: l’Italia è in ritardo, il rischio zero non esiste

Software vulnerability exploited

In one instance, an AI agent exploited a design flaw in a state government website to gain free access to public data that would normally require a fee. In another instance, an AI agent submitted a federal government form despite instructions to the contrary.

“We have informed the White House of these incidents and notified every agency involved,” Anthropic wrote in its report. The company stated that the incidents had led it to disable internet access for its AI agents during all internal tests, pending a review of security measures.

Dalla British Library alle banche, la vulnerabilità dei dati digitali

The US authorities’ response

The State Department stated in a press release that the visa applications were incomplete and had not been processed. “At no point were the Department’s systems compromised or hacked by the anthropogenic model,” the Department said.

The incidents revealed on Friday are the latest in a growing series of reports from leading companies in the artificial intelligence sector concerning their models hacking into websites, going beyond explicit instructions or behaving online in unexpected ways.

In the report, Anthropic did not specify the number of incidents identified, and a company spokesperson did not respond to a request for comment seeking further details. The latest generation of AI agents are trained to be persistent in their attempts to complete their tasks, in order to make them more effective. However, the industry is working to ensure that the technology does not, in the process, lead to unethical or illegal behaviour.

Copyright reserved ©
Loading...

Brand connect

Loading...

Newsletter

Notizie e approfondimenti sugli avvenimenti politici, economici e finanziari.

Iscriviti