AI Dynamics

Global AI News Aggregator

About

ARC-AGI-3 Benchmarks Agent Performance Against Human Action Efficiency

ARC-AGI-3 scores agents on how close they are to human action efficiency. All ARC-AGI-3 environments were solved by at least 2 human testers out of 10 (most of the time it was 5+). We use the action count of the 2nd best tester (to avoid outlier performance) as our human

→ View original post on X — @fchollet