97% OF PERMISSION PROMPTS IN CLAUDE CODE GET APPROVED WITHOUT BEING READ
a coworker sent me the number with no comment. that was enough to make it unsettling.
anthropic admitted this on august 7, right when they announced it.
but that’s not even the interesting part.
they ran a controlled study with 1,053 paid testers.
they planted a dangerous command inside a task. humans caught it 13.6% of the time. the auto mode classifier caught it 89% of the time.
starting august 14, auto mode becomes the default for new sessions on pro, max, and team plans.
it only blocks actions that are irreversible, destructive, or reach outside your environment. everything else goes through without stopping.
not because it’s faster. because it turned out to be safer than a human clicking “approve” on autopilot.
so you were clicking “yes” without thinking, and now a model does the thinking instead. and it’s better at it.
显示更多