AI Dynamics

Global AI News Aggregator

About

Micro-benchmarks Don’t Measure True Reasoning Capabilities

These micro-benchmarks are fun but I've found the model that "wins" changes depending on the exact framing of the prompt. Pattern matching != reasoning.

→ View original post on X — @whats_ai