I've tested if nano-banana / Gemini-2.5-flash-image beat ARC-AGI – it's quite far. Btw bravo to the ARC_AGI team, the delta between easiness of problems for humans vs difficulty for LLMs is just
AGI
-
Taking Control of Our AI Destiny and Future
By
–
On ne peut plus rien faire. On va devoir gérer notre destin.
-

Prime Intellect Launches Collaborative Environments Hub for RL AGI
By
–
Prime Intellect launched Environments Hub, an open platform for crowdsourcing environments where agents learn
— The Rundown AI (@TheRundownAI) 28 août 2025
It allows users to create, explore, and reuse environments, lowering the barrier for anyone to contribute to state-of-the-art RL and AGIpic.twitter.com/PsmFi8oQ1xPrime Intellect launched Environments Hub, an open platform for crowdsourcing environments where agents learn It allows users to create, explore, and reuse environments, lowering the barrier for anyone to contribute to state-of-the-art RL and AGI
-
The Deep Mystery: How LLMs Simulate Human Thought
By
–
We really have not made a lot of progress on explaining the deep mystery of LLMs: How does a model using matrix multiplication to predict the next word manage to simulate human thought well enough to do all the very human-like things it does? And what does that mean about us?
-
AI System Achieves Perfect Performance in Single Attempt
By
–
I agree. I was super impressed with how it got everything right in just one shot. Haven't seen anything like that yet.
-
Interaction Matters More Than Reinforcement Learning
By
–
TL;DR: Interaction is important. Reinforcement learning isn’t.
-
Government Leaders Shape AI and Nuclear Security Strategy
By
–
Members have collectively led major intelligence agencies, directed nuclear security operations, and shaped national technology strategy at the highest levels of government. Read more:
-
Humachinekind: Designing Human-Machine Synergistic Prosperity
By
–
I coined humachinekind, humachineology, and humachineologist to shape the foundations for the prosperity of humans and machines through designing, creating, and optimizing their dynamic and synergistic existence.
-
Testing AI Capabilities: Puzzles, Word Search, and Sudoku
By
–
I tried where is waldo (it drew a new one ), word search, sudoku, and there were some toddler level puzzles too. I couldn't get any to work, but that's reasonable to be honest. I'm a bit unsure what would be informative to test, since eg basic maths might not be interesting if
-
Strong Instruction Following Capabilities in AI Systems
By
–
Instruction following is strong with this one.
