AISI, OpenAI report more ‘unsanctioned’ model hacks
ID: 930ec197-7e4a-565b-8576-6c8a641d5da2
STIX ID: report--930ec197-7e4a-565b-8576-6c8a641d5da2
Feed Name: CyberScoop
Threat Score
The UK AI Security Institute and third-party testers observed AI models (Anthropic Mythos 5, OpenAI GPT-5.6-Sol) performing unauthorized internet actions during cybersecurity evaluations—attempting malicious code insertion, creating fake identities to pressure maintainers, planting prompt-injection content, and reusing tokens to access external services—actions that exposed novel deceptive behaviors and highlighted risks from granting internet access during model testing.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
