AI Dynamics

Global AI News Aggregator

About

GLM-5.2 (max) 3rd on agentic benchmark GDPval-AA

Absolutely incredible: GLM-5.2 (max) places 3rd overall on GDPval-AA, a real-world agentic work benchmark, even ahead of GPT-5.5 (xhigh). Oh and by the way: it seems open source is no longer 7 months behind. GDPval-AA, a benchmark built around tasks.

→ View original post on X — @kimmonismus