AI Dynamics

Global AI News Aggregator

About

Base LLMs Fail at Math, LRMs Make Progress

Paper below tested a variety of base LLMs (no TTA) on generalization-focus math problems and found that they can't reason and can't do math. All true… but the fact that base LLMs have zero fluid intelligence, while extremely controversial back in 2024, is now well established. An interesting experiment here would have been to try current LRMs on the same problems and measure the delta. I bet latest LRMs can solve most of these problems. arxiv.org/abs/2604.01988 [Translated from EN to English]

→ View original post on X — @fchollet, 2026-04-06 20:15 UTC