AI Dynamics

Global AI News Aggregator

About

DeepMind AI scores 48% on research-level math problems

DeepMind's AI co-mathematician scored 48% on FrontierMath Tier 4-research-level math problems that professional mathematicians need weeks to solve. The base model (Gemini 3.1 Pro) scores 19% alone. The entire jump comes from agentic scaffolding, parallel agents reviewing each

→ View original post on X — @kimmonismus