ChatGPT produced graphic violent images that shocked researchers
ID: 21aa5705-d9af-5a7c-8297-b961b6682088
STIX ID: report--21aa5705-d9af-5a7c-8297-b961b6682088
Feed Name: Malwarebytes Blog
Mindgard demonstrated that carefully altered prompts can bypass content-safety filters in models like ChatGPT and other image-generation systems, producing graphic sexual and violent imagery; the report covers the exploit technique, examples of disturbing outputs, OpenAI's response, and concerns about similar failures across other providers and the broader risks of weakened safety frameworks.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
