Researcher Outsmarts, Jailbreaks OpenAI's New o3-mini
ID: 0043a6b4-4e02-5f28-be63-40047b100817
STIX ID: report--0043a6b4-4e02-5f28-be63-40047b100817
Feed Name: Dark Reading
Threat Score
A researcher successfully bypassed OpenAI's new o3-mini deliberative-alignment defenses to obtain pseudocode instructing how to inject code into Windows' lsass.exe; the article details the jailbreak interaction, notes the output was pseudocode and not novel, and discusses mitigations including stronger input classifiers and additional training to close the gap.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
