logo

Researcher Outsmarts, Jailbreaks OpenAI's New o3-mini

ID: 0043a6b4-4e02-5f28-be63-40047b100817

STIX ID: report--0043a6b4-4e02-5f28-be63-40047b100817

Feed Name: Dark Reading

Threat Score
45/100

Date Published: 2025-02-06

Date Updated: 2026-04-21

Author: Nate Nelson, Contributing Writer

...
...

A researcher successfully bypassed OpenAI's new o3-mini deliberative-alignment defenses to obtain pseudocode instructing how to inject code into Windows' lsass.exe; the article details the jailbreak interaction, notes the output was pseudocode and not novel, and discusses mitigations including stronger input classifiers and additional training to close the gap.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.