this is keynote worthy and means even more coming from a coding agent creator. i am always pro thoughtfulness and we love featuring calls for sanity amongst the hype (my agents conf opened with @sayashk pointing out how most agents are overhyped and it was so well received)
AI
-

AGI Benchmarks, SpaceX IPO, Meta’s Trillion Vision, Sora Shutdown
By
–
Benchmark's Future, ARC-AGI, SpaceX IPO, Epic Games Layoffs, Meta Aims for $9 Trillion, RIP Sora https://
x.com/i/broadcasts/1
rGmqoDVwbZGy
… -

ARC-AGI-3 Benchmarks Agent Performance Against Human Action Efficiency
By
–
ARC-AGI-3 scores agents on how close they are to human action efficiency. All ARC-AGI-3 environments were solved by at least 2 human testers out of 10 (most of the time it was 5+). We use the action count of the 2nd best tester (to avoid outlier performance) as our human
-
Claude AI Ships New Features at Remarkable Pace
By
–
When was the last day that Claude DIDN’T ship a new feature? https://t.co/pOY0gvgemQ
— Matt Wolfe (@mreflow) 25 mars 2026When was the last day that Claude DIDN’T ship a new feature?
-
Frontier AI Achievement: Low Probability Assessment and Reasoning
By
–
The topics I brought up were more controversial than I thought, but it's reaching the right people and the recent Composer 2 tech report released since is encouraging too. My underlying opinion has not changed though. Here's my reasoning: – Low chance of achieving frontier:
-
AGI Capability: Task Learning Without Human Intervention Required
By
–
Either you believe AGI is possible, in which case a real AGI will be able to look at ARC-AGI-3 and ace it, because regular humans can… …or you believe that AI is just an automation tool that will require human intervention every time a new task comes up. Pick your camp.
-

ARC-AGI 3 Highlights Remaining Gaps Between Human and AI Capabilities
By
–
Veremos como acabamos el año! Como siempre, ARC-AGI 3 encontrando esos gaps donde humanos e IA se diferencian ¿Resolverlo nos ofrece una AGI? No, pero la existencia de estos gaps de capacidades nos indican que aún queda mucho trabajo por hacer.
-
Google Launches Lyria 3 Pro Advanced AI Music Generation Model
By
–
Just a month after dropping Lyria 3, @Google is already turning up the volume! 🎛️
— Harold Sinnott #MWC26 (@HaroldSinnott) 25 mars 2026
They just announced Lyria 3 Pro from @GoogleDeepMind , their most advanced AI music model yet. Creators can now build tracks up to 3 minutes long with way more creative control. Plus, they’re… pic.twitter.com/RzG5fm5hjEJust a month after dropping Lyria 3, @Google is already turning up the volume! They just announced Lyria 3 Pro from @GoogleDeepMind , their most advanced AI music model yet. Creators can now build tracks up to 3 minutes long with way more creative control. Plus, they’re
-

What General Intelligence Really Means for AGI
By
–
The G in AGI stands for "general". General intelligence does not mean that you have been specifically trained for a large range of tasks. It means you can approach any NEW task and figure it out, just like humans do. If regular people can do it on their own (no guidance, no
-

The AI Scientist Published in Nature: Fully Automated Research
By
–
The AI Scientist: Towards Fully Automated AI Research, Now Published in Nature!!✨ Today in Nature we share a comprehensive technical summary of our work on The AI Scientist, including new scaling law results showing how it improves with more compute and more intelligent foundation models. The AI Scientist autonomously creates its own research ideas, codes up and conducts experiments to test those ideas, creates figures to visualize the results, writes an entire scientific manuscript summarizing what it has discovered, and conducts its own “peer” review of the resulting paper. One of its papers–entirely AI generated–passed peer review at a top-tier AI conference workshop, a historic milestone marking the dawn of a new era of AI-accelerated scientific discovery. 🔬🧪✨🧬💡🔭 Paper nature.com/articles/s41586-0… Blog sakana.ai/ai-scientist-natur… Work done in collaboration with a great team from Sakana, Oxford, and my lab at UBC. Thanks and congratulations everyone! @_chris_lu_ @cong_ml @RobertTLange @_yutaroyamada @shengranhu @j_foerst @hardmaru
→ View original post on X — @_yutaroyamada, 2026-03-25 18:01 UTC