“An Anthropic study found that AutoMode in Claude code caught 89% of harmful actions, while human reviewers only caught 13.6%.”