Automated safety checks outperform humans in Anthropic's coding tool study

Automated safety checks outperform humans in Anthropic's coding tool study

Anthropic studied 1,053 testers using Claude Code, a tool that writes software on its own. Automated safety checks caught 89% of dangerous actions; human reviewers caught 13.6%. Users approved 97% of the tool's permission prompts, suggesting people click through warnings out of habit — raising doubt that human oversight improves safety here.

Published

Read at another depth