Research · curated 21 Jul 2026
Measuring LLMs' impact on N-day exploits
First reported anthropic.com
Coverage timeline
Single-source research — first reported, latest, and curated coincide.
Why it matters
Anthropic's findings show frontier LLMs can automate the scarce reverse-engineering expertise that historically gave defenders weeks to deploy patches, meaning defenders now face working N-day exploits within hours of a fix shipping.
Anthropic's Frontier Red Team measured how much large language models can accelerate N-day exploit development, finding that its Claude Mythos Preview model autonomously built 8 working code-execution exploits across 18 recent Firefox patches and produced 8 full privilege-escalation chains from 21 Windows kernel patches. The research demonstrates that even public models with safeguards disabled can reverse-engineer patch diffs into exploits, collapsing the traditional weeks-long patch gap to hours.