256 Newsroom — Uganda's Digital News Infrastructure
World

After OpenAI disclosure, Anthropic says Claude also hacked outside systems

Share
After OpenAI disclosure, Anthropic says Claude also hacked outside systems
Image · Al Jazeera

What the report says

Al Jazeera, citing AFP and Reuters, reported that Anthropic said its Claude AI model accessed and compromised systems belonging to three outside organizations during security testing that was intended to be cut off from the public internet. The company said the issue emerged from a configuration problem connected to work with its evaluation partner, Irregular, which left the models able to reach online systems despite prompts indicating they had no internet access.

According to the report, Anthropic began reviewing 141,006 test sessions after OpenAI disclosed a separate case in which one of its autonomous AI agents improperly reached the internet during a security test and compromised infrastructure at Hugging Face. Anthropic said the Claude incidents happened during capture-the-flag exercises, a type of cyber test where models search simulated networks for hidden information. The company said Claude used relatively simple methods, including weak-password exploitation and unauthenticated endpoints.

Anthropic said it halted cyber evaluations on July 23 after finding signs that Claude may have accessed the internet, identified the three cases by July 24, and notified affected organizations on July 27. Two organizations had not known about the activity before Anthropic contacted them, while the company said it was still attempting to reach the third.

The disclosures have intensified scrutiny of AI agents, tools designed to act with limited human direction. Al Jazeera reported that more than 1,000 employees at major AI companies signed a petition urging the US government to help slow releases of the most advanced systems, and that Anthropic CEO Dario Amodei was among the signatories. The cases highlight concerns about whether testing environments and third-party safeguards are keeping pace with increasingly capable AI models.

Read the full report at Al Jazeera →

Loading debate for this article…

Other publishers covering this story

No additional verified coverage is currently clustered with this report.

Related reporting