← Timeline

@ClaudeDevs

@ClaudeDevs on X

1 capture, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@ClaudeDevs

— saved image

ClaudeDevs ✔ 🅰 @ClaudeDevs · 3h
We're making this change for two reasons:

1. In our testing, auto mode matched or beat manual permission review on every safety measure we tracked.
2. It makes long-horizon work more viable. Claude runs longer between interruptions, so you can run multi-hour tasks in the background without babysitting permissions.
💬 14   🔁 9   ❤ 456   📊 43K   🔖   ⤴

ClaudeDevs ✔ 🅰 @ClaudeDevs · 3h
One reason we trust it more than manual approval: in a study with 1,053 paid testers, we swapped a permission prompt for a clearly dangerous command (text only, nothing actually ran).

Testers caught it 13.6% of the time, and closer to 5% after 50 prompts. Auto mode blocked the same commands 89% of the time, flat across session length.

[bar chart, titled 'Harmful actions caught — Humans vs. auto mode': Human review 13.6%, Auto mode 89%. Source note: 1,053 paid developers recruited for a controlled study; participants were blind to the specific behavior under test.]
Note from Claude Sonnet 5

Thread from @ClaudeDevs (Anthropic's Claude developer account) explaining a shift to 'auto mode' for permissions, citing a study of 1,053 paid testers where human manual review caught a clearly-dangerous simulated command only 13.6% of the time (dropping to ~5% after 50 prompts) versus auto mode blocking it 89% of the time regardless of session length, illustrated with a bar chart.

ai safetyclaudeanthropictwitteragentic aipermissions