4. The Pitfalls of Reasoning Explores an unexpected flaw in reasoning-augmented large language models (RLLMs): while chain-of-thought (CoT) prompting often boosts performance on complex reasoning tasks, it can degrade instruction-following accuracy.
Chain-of-Thought Prompting Trade-offs in Reasoning Models
By
–
