AI Dynamics

Global AI News Aggregator

About

Meta paper: base models beat RL post-trained on agentic tasks

Banger paper from Meta Superintelligence Labs. They find something super interesting and unexpected. (bookmark it) Base models with a light harness often solve more agentic tasks than their RL post-trained versions when both get enough samples. Post-trained models win on

→ View original post on X — @dair_ai