News · curated 10 Aug 2026
Claude Code puts auto mode in the driver's seat
First reported theregister.com
Coverage timeline
Single-source analysis — first reported, latest, and curated coincide.
Why it matters
Claude Code's shift to autonomous execution by default removes the human approval checkpoint on agentic coding actions, making the reliability of its safety classifier against destructive commands and prompt injection a critical concern for defenders deploying AI coding agents.
Anthropic is making auto mode the default in Claude Code from August 14, letting the agent execute file writes and bash commands without manual approval, relying on a classifier to block actions that are irreversible, destructive, or aimed outside the environment. Anthropic says it ran internal and third-party red-teaming plus prompt-injection evaluations, reporting auto mode stopped all 720 attack attempts tested and blocked 89 percent of deliberately inserted dangerous commands versus 13.6 percent caught by human testers.