logo

ChatGPT, Claude, and Gemini Among 11 AI Models Vulnerable to One-Line Jailbreak

ID: 68c4f4ba-dec1-5cdf-a38d-e760fdc37634

STIX ID: report--68c4f4ba-dec1-5cdf-a38d-e760fdc37634

Feed Name: GBHackers

Threat Score
70/100

Date Published: 2026-04-10

Date Updated: 2026-04-22

Author: Divya

...
...

A newly reported "sockpuppeting" jailbreak exploits assistant-prefill (prefix injection) at the API layer to cause LLMs to continue generating restricted outputs by inserting a fake acceptance message; tested across 11 models, it enabled generation of malicious code and system-prompt leakage in some cases, and can be mitigated by enforcing message-order validation at the API layer though self-hosted inference servers often remain vulnerable.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.