There's going to be a lot more software, and a lot more demand for software engineers. And a lot more token consumption.
@fchollet
-
ARC-AGI-4 Benchmark Release Planned Early 2027
By
–
For those wondering about ARC-AGI-4 timing: it will be released in early 2027. We are aiming for a yearly release schedule for new benchmarks. We are also aiming for each new benchmark to be fully unsaturated upon release, and to target the most important unanswered research
-
ARC Prize Sets Actionable Goals for AGI Research Progress
By
–
ARC Prize's mission is to provide actionable goals for AGI research and to measure progress towards them.
-
AGI Autonomy: Self-Adaptation for New Tasks
By
–
Yes, if it's AGI it should be able to "make its own harness" for a new task, or just internalize it.
-
AGI Definition Unchanged Since 2019 Compass Not Target
By
–
I have been saying the same thing for years. Initially I was using this exact line ("it's a compass, not a target to hit") about ARC 1, back around 2021-2022, before ChatGPT. This has always been our stance. My bar for AGI has been public and unchanged since 2019. It's AGI when
-
AI Model Development Pace Challenges Environment Creation Timeline
By
–
This would have been materially impossible to do even if we wanted to, since it took us an entire year to create the environments, while new AI models come out every month. By the time the latest SotA model is out, the full set of environments already is already shipped.
-
ARC-AGI: Measuring Progress Toward Artificial General Intelligence
By
–
Keep in mind: ARC-AGI is *not* a final exam that you pass to claim AGI. Including ARC-AGI-3. The benchmarks target the residual gap between what's hard for AI and what's easy for humans. It's meant to be a tool to measure AGI progress and to drive researchers towards the most
-
Human-Level General Intelligence Requires Learning Efficiency Parity
By
–
Human-level general intelligence is achieved when an AI system can approach a new task and figure it out, without human intervention, *with the same learning efficiency as humans*. If every new task requires human intervention, it's not general. If every new task requires
-
ARC-AGI-3 Game Studio Launch with Exceptional Team
By
–
To ship ARC-AGI-3, we've set up an entire game studio. Exceptional crew — super proud of the team. They did an incredible job.
-
ARC-AGI-3 Leaderboard: Measuring True AI Intelligence Beyond Human Design
By
–
It is trivial to solve all public ARC-AGI-3 tasks if you have a human looking at them and designing a system to beat them (we have released a harness that uses human replay to score 100%). But our leaderboard is not about measuring how well human intelligence does on ARC-AGI-3,