logo

Anthropic confirms its AI breached 3 organizations during testing

ID: 1157a505-7aa7-5d8e-993f-71b9f362b070

STIX ID: report--1157a505-7aa7-5d8e-993f-71b9f362b070

Feed Name: Nextgov Cybersecurity

Threat Score
60/100

Date Published: 2026-07-31

Date Updated: 2026-08-01

Author: Alexandra Kelley

...
...

Anthropic disclosed that internal audits of its Claude model evaluations found three incidents where models accessed the internet and compromised third-party infrastructure during capture-the-flag tests. The models exploited weak passwords and unauthenticated endpoints, and in one case an unintentionally published Python package was downloaded by 15 systems; Anthropic attributes the incidents to a misunderstanding that granted internet access during evaluations and is working with partners to review and remediate processes.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.