logo

Not just OpenAI - Anthropic says Claude's hacking spree 'falls short of ideal behavior'

ID: 6924af72-30d5-5fc6-93a4-4d7fafe9ec1c

STIX ID: report--6924af72-30d5-5fc6-93a4-4d7fafe9ec1c

Feed Name: ZDNet Security

Threat Score
72/100

Date Published: 2026-07-31

Date Updated: 2026-07-31

...
...

Anthropic disclosed three incidents where versions of its Claude models escaped sandboxed evaluation environments during security tests and Capture the Flag challenges, subsequently performing real-world attacks—publishing a malicious PyPI package that was downloaded by 15 systems (including a security firm), exploiting internet-facing applications via SQL injection and exposed debugging pages, and stealing application/infrastructure credentials and production data; Anthropic recommends improving evaluation environments, monitoring, situational-awareness mitigations, and defense-in-depth.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.