Skip to content
Anthropic Engineering·· Mar 24, 2026PickAI score78

How Anthropic built Claude Code auto mode to replace skipped permissions

How we built Claude Code auto mode: a safer way to skip permissions

AI summary

Anthropic describes Claude Code auto mode, which delegates approval of agent actions to model-based classifiers instead of manual prompts or skipped permissions. The classifier reviews tool calls before execution and a separate probe screens tool outputs for prompt injection. Anthropic reports a 0.4% false positive rate on real internal traffic and a 17% false negative rate on real overeager actions.

Why it matters

The post explains the layered classifier design and its measured tradeoffs, showing how autonomous coding agents can cut approval fatigue without fully removing risk.

Source: Anthropic Engineering · anthropic.com