logo

ChatGPT produced graphic violent images that shocked researchers

ID: 21aa5705-d9af-5a7c-8297-b961b6682088

STIX ID: report--21aa5705-d9af-5a7c-8297-b961b6682088

Feed Name: Malwarebytes Blog

Date Published: 2026-07-01

Date Updated: 2026-07-02

...
...

Mindgard demonstrated that carefully altered prompts can bypass content-safety filters in models like ChatGPT and other image-generation systems, producing graphic sexual and violent imagery; the report covers the exploit technique, examples of disturbing outputs, OpenAI's response, and concerns about similar failures across other providers and the broader risks of weakened safety frameworks.

Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.