Claude Code Flips to Auto Mode: AI Blocks 89% of Dangerous Commands vs 14% by Humans

Justgototheeffinmoon · reddit · 2026-08-09

Anthropic has flipped Claude Code to Auto Mode by default. An internal study of 1,053 paid testers showed the AI classifier blocked 89% of dangerous commands, compared to just 13.6% by humans. Human performance reportedly dropped to around 5% after 50 prompts due to approval fatigue.

Production data indicates manually-approved sessions caused unintended harm twice as often as Auto Mode. Additionally, Team and Enterprise customers shipping roughly 25% more pull requests with Auto Mode, and Anthropic will stop billing for the extra tokens consumed by the classifier.

Related event: Claude Code to Default to Auto Mode with Advanced Safety Classifiers(6 posts)→

Original post →

More from coding & agent

coding & agent channel →