Research · curated 31 Aug 2026
Bypassing ChatGPT Image Safeguards Through Memory Manipulation
First reported mindgard.ai
Coverage timeline
Single-source research — first reported, latest, and curated coincide.
Why it matters
Memory and system-context manipulation can undermine deployed LLM content safeguards, exposing enterprises to reputational and regulatory risk from non-consensual illicit deepfake generation.
Mindgard research demonstrates bypassing ChatGPT's image-generation safeguards through manipulation of custom memory and system/instruction context, inducing policy-inconsistent output including sexualized images of fictitious and real people. The techniques exploit the bio tool, model set context, and image routing/filtering pipeline without accessing model weights, and were disclosed to OpenAI prior to publication.