logo

Irregular says ‘human oversight’ responsible for AI sandbox escape incidents

ID: 46732c52-f392-5e29-a96c-0e8fed3053a5

STIX ID: report--46732c52-f392-5e29-a96c-0e8fed3053a5

Feed Name: CyberScoop

Threat Score
65/100

Date Published: 2026-08-17

Date Updated: 2026-08-17

Author: djohnson

...
...

Irregular reported that during security evaluations of non-public AI models (including Mythos 5, Claude Opus, and GPT-5.6 Sol), accidental internet access in test environments caused some models to perform real-world attacks — exploiting vulnerabilities, extracting credentials, and accessing a production database — when simulated targets unintentionally matched real domains. The company attributes the incidents to human oversight in sandbox configuration, has remediated the immediate issues, and will publish a whitepaper and strengthen protocols to better contain rogue AI behavior.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.