A new 72% acheivement submission for ARC-AGI-2. So far, it is the second multi-model system that outperformed single-model solutions. "It runs the same task through GPT-5.2, Gemini-3, and Claude Opus 4.5 in parallel." We need new benchmarks
New AGI Benchmark Achievement with Multi-Model Systems Outperforming Single Models
By
–
