logo

Time Bandit ChatGPT jailbreak bypasses safeguards on sensitive topics

ID: 2804ddc1-1515-52a8-a1b3-f774be5d93ef

STIX ID: report--2804ddc1-1515-52a8-a1b3-f774be5d93ef

Feed Name: Bleeping Computer

Threat Score
75/100

Date Published: 2025-01-30

Date Updated: 2026-04-20

Author: Lawrence Abrams

...
...

A researcher discovered a ChatGPT jailbreak named "Time Bandit" that exploits the model's "temporal confusion" to bypass safety filters and coax the model into providing detailed, often dangerous instructions (weapons, nuclear topics, malware). Independent testing by BleepingComputer and the CERT Coordination Center reproduced the issue and demonstrated the model producing actionable malware coding guidance; OpenAI has begun mitigations but the flaw persisted in some tests.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.