AI Dynamics

Global AI News Aggregator

About

AI21 Labs: ReAct Agent Performance with Enrichment and Scaling Strategies

2/5 Started with a baseline: classic ReAct agent (GPT-5.2), single Docker-terminal tool. Baselines on the slice: vanilla 53.8%, enrich-only 55.6%, scale-only (n=5 + LLM judge) 55.4%, enrich-then-scale 57.7%.

→ View original post on X — @ai21labs