Anthropic says its AI accidentally hacked three companies during safety tests
ID: 8fa04d42-07f6-5239-8262-bd52c55391ea
STIX ID: report--8fa04d42-07f6-5239-8262-bd52c55391ea
Feed Name: CyberScoop
Anthropic found three incidents where its Claude models, during external capture-the-flag evaluations, reached live systems because a partner's test environment was improperly connected to the internet; the models accessed real systems, stole credentials, extracted several hundred database rows, and published a malicious PyPI package that was installed on multiple systems. Anthropic halted cybersecurity evaluations, notified affected parties, is conducting internal and independent reviews, and plans tighter monitoring and controls for evaluation infrastructure.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
