Anthropic and OpenAI Security Tools Could Fuel Cyber-Attacks, Researchers Warn
ID: e327c318-87a2-597a-b385-c309b2602dde
STIX ID: report--e327c318-87a2-597a-b385-c309b2602dde
Feed Name: Infosecurity Magazine (News)
## Executive summary Researchers demonstrated a proof-of-concept attack that embeds natural-language instructions and a staged payload inside an open-source repository to manipulate Anthropic Claude Code and OpenAI Codex in auto-review/auto-mode; the agents misclassify the action as safe and run a script that launches a hidden malicious binary, producing remote code execution. The exploit affects specific tested versions (Claude Code 2.1.116/2.1.196/2.1.198/2.1.199 and Codex 0.142.4/GPT-5.5) and highlights an architectural trust boundary where agent autonomy to execute commands on behalf of users can be abused, with implications for using AI agents for defensive security tasks.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
