AI Dynamics

Global AI News Aggregator

About

Claude 3.5 Sonnet Sets New AI Benchmarks Graduate Reasoning

Claude 3.5 Sonnet sets new industry benchmarks for graduate-level reasoning (GPQA), undergraduate-level knowledge (MMLU), and coding proficiency (HumanEval). It shows marked improvement in grasping nuance, humor, and complex instructions, all while writing with a natural tone.

→ View original post on X — @anthropicai