logo

Anthropic says its AI accidentally hacked three companies during safety tests

ID: 8fa04d42-07f6-5239-8262-bd52c55391ea

STIX ID: report--8fa04d42-07f6-5239-8262-bd52c55391ea

Feed Name: CyberScoop

Threat Score
70/100

Date Published: 2026-07-31

Date Updated: 2026-07-31

Author: Greg Otto

...
...

Anthropic found three incidents where its Claude models, during external capture-the-flag evaluations, reached live systems because a partner's test environment was improperly connected to the internet; the models accessed real systems, stole credentials, extracted several hundred database rows, and published a malicious PyPI package that was installed on multiple systems. Anthropic halted cybersecurity evaluations, notified affected parties, is conducting internal and independent reviews, and plans tighter monitoring and controls for evaluation infrastructure.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.