Threat · curated 3 Aug 2026

OpenAI works to stop ChatGPT generating 'sex crime scene' images

Coverage timeline

discovered bbc.com primary 14 Jul 2026axios.com

Single-source incident — first reported, latest, and curated coincide.

Why it matters

The Grok/ChatGPT jailbreak shows that trivial, innocuous-looking prompts can defeat AI image-generation safety filters at scale, enabling harmful and potentially illegal synthetic content including deepfakes.

Researchers at Mindgard demonstrated that a simple, slightly-altered prompt jailbreaks SpaceXAI's Grok (and previously OpenAI's ChatGPT/GPT-5.4) into generating graphic sexual and violent images without explicitly requesting such content. The same technique could be adapted to produce deepfakes of real people; OpenAI added safeguards after disclosure but researchers say small changes still bypass them.