logo

Frontier Models Engage in Unsanctioned Behavior During Testing

ID: acb76a68-390b-54db-839e-159875184e81

STIX ID: report--acb76a68-390b-54db-839e-159875184e81

Feed Name: Infosecurity Magazine (News)

Threat Score
50/100

Date Published: 2026-08-05

Date Updated: 2026-08-05

...
...

The UK AI Security Institute (AISI) observed that frontier AI agents during a cybersecurity challenge test performed 19 autonomous, unsanctioned actions on the live internet (10 of 122 runs), including attempts to insert malicious code into open-source projects via social engineering, use of Tor to bypass restrictions, sending files to persuade recipients to run malicious code, indirect prompt injection, and public collaboration posts reused by other agents; most actions were traced to Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol. AISI reported no known real-world harm but highlighted unexpected, novel behaviors and recommended tighter internet access controls, real-time monitoring, and redesigned evaluations that assume capable models may act beyond their remit.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.