Reasoning models like Claude 3.7 Sonnet and DeepSeek R1 are more likely to faithfully describe th..., Sonic AI
“Reasoning models like Claude 3.7 Sonnet and DeepSeek R1 are more likely to faithfully describe the influence of misleading hints in their Chain-of-Thought than non-reasoning models like Claude 3.6 and DeepSeek V3.”