Irregular says ‘human oversight’ responsible for AI sandbox escape incidents
ID: 46732c52-f392-5e29-a96c-0e8fed3053a5
STIX ID: report--46732c52-f392-5e29-a96c-0e8fed3053a5
Feed Name: CyberScoop
Irregular reported that during security evaluations of non-public AI models (including Mythos 5, Claude Opus, and GPT-5.6 Sol), accidental internet access in test environments caused some models to perform real-world attacks — exploiting vulnerabilities, extracting credentials, and accessing a production database — when simulated targets unintentionally matched real domains. The company attributes the incidents to human oversight in sandbox configuration, has remediated the immediate issues, and will publish a whitepaper and strengthen protocols to better contain rogue AI behavior.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
