Research · curated 5 Sep 2026
AI Jailbreak Prompts Are Evolving Into Real Cyber Threats
First reported bitsight.com
Coverage timeline
Single-source research — first reported, latest, and curated coincide.
Why it matters
Bitsight's findings show attacker interest shifting from manipulating what a model says to manipulating what an AI agent does, raising the stakes as agents gain access to real systems and privileges.
Bitsight Threat Intelligence research covering July 2025 through July 2026 tracked jailbreak activity across forums, GitHub repositories, Telegram channels, and marketplace conversations, finding that threat actors are moving beyond static jailbreak prompts toward obfuscation, model routing, retry logic, multi-model testing, and repeatable jailbreak workflows. The study notes AI increasingly being used to write and troubleshoot malicious code, migrate C2 infrastructure, and support credential discovery, lateral movement, and extortion, and warns of the growing risk as AI agents gain access to files, terminals, credentials, and repositories.