ICYMI: I interviewed the @theaidocfilm
's creators during @sxsw and wrote about it for @FastCompany
. It's a documentary everyone should see, and *especially* if you don't work in AI. Yes, it's a film about the promise, peril, and uncertainty surrounding AI. Yes, it features
RESEARCH
-
AI Documentary Explores Promise Peril and Uncertainty in Artificial Intelligence
By
–
-
Blue Brain Project Failed Due to Wrong Neuroscience Hypothesis
By
–
I think the Blue Brain project failed because the main underlying hypothesis about how the brain works may be wrong: simulating textbook neuroscience theories does not produce human-like output. But transistor based systems already do!
-

Adversarial Examples and the Challenge of Exponential Misalignment
By
–
"Solving adversarial examples requires solving exponential misalignment" This paper argues that adversarial examples may happen because neural nets treat “cat”, “dog”, etc. as huge regions of image space, including many weird, non-human-looking inputs, so tiny edits can flip
-

OpenAI’s AI Intern Coming September 2026, Full System by 2028
By
–
OpenAI's first “AI intern” expected by September and a full system targeted for 2028. Powered by advances in reasoning models and agent systems like Codex, these tools already show dramatic productivity gains, solving problems in days instead of weeks, but still face
-
AI Satellite Mapping Tackles Global Health Crisis Affecting 250M
By
–
250M people affected, mostly kids. This is exactly where AI should be going. From dissecting snails to satellite mapping is next level.
-
Dream2Flow: 3D Object Flow for Robot Manipulation from Video
By
–
Our recent work using object-centered spatial information for better generalization 🦾 https://t.co/aWFfOwxSov
— Fei-Fei Li (@drfeifei) 20 mars 2026Our recent work using object-centered spatial information for better generalization 🦾 Wenlong Huang (@wenlong_huang) What representation enables open-world robot manipulation from generated videos? Introducing Dream2Flow, our recent work that bridges video generation and robot control with 3D object flow. dream2flow.github.io @Stanford #ICRA2026 1/N — https://nitter.net/wenlong_huang/status/2035032566529712244#m
-
AI-Empowered Mathematicians Achieve H2 Level Research Work
By
–
And shoutout to this independent work by @JulianSlzr and team on a Level 2 work, in the "Primarily Human" category, with the help of AlphaEvolve and DeepThink! nitter.net/JulianSlzr/status/2034… Julian Salazar (@JulianSlzr) We're AI researchers @GoogleDeepMind who last did math full-time over 9 years ago. Despite our rustiness and limited time, AI empowered us to do some niche theory-building (@littmath). Per Aletheia's taxonomy, our work is H2 (primarily human)… for now! nitter.net/lmthang/status/2021644… — https://nitter.net/JulianSlzr/status/2034947452005228627#m
-

AI Accelerates Mathematical and Scientific Discovery with Gemini DeepThink
By
–
And before that 6 other math research papers by #Aletheia and other discoveries in physics and computer science, all of which were powered by Gemini #DeepThink! nitter.net/lmthang/status/2021631… Thang Luong (@lmthang) 6 months in, after the IMO-gold achievement, I’m very excited to share another important milestone: AI can help accelerate knowledge discovery in mathematics, physics, and computer science! We’re sharing Two new papers from @GoogleDeepMind and @GoogleResearch that explore how Gemini #DeepThink together with agentic workflows can empower mathematicians and scientists to tackle professional research problems. Some highlights: The first paper built a research agent #Aletheia, powered by an advanced version of Gemini Deep Think, that can autonomously produce publishable math research and crack open Erdős problems. The second paper, built on similar agentic reasoning ideas, helped resolve bottlenecks in 18 research problems, across algorithms, ML and combinatorial optimization, information theory and economics. See the thread for details about the two papers and the joint blog post. — https://nitter.net/lmthang/status/2021631397614731563#m
-

Aletheia solves 7th FirstProof problem at publishable research level
By
–
Tackling FirstProof was our 7th math research paper, which was done autonomously and the solution to problem #7 is at Level 2 "Publishable Research" as well. nitter.net/lmthang/status/2026689… Thang Luong (@lmthang) Thrilled to share: #Aletheia, our math research agent, just solved 6/10 notoriously hard FirstProof problems autonomously, the best result in the inaugural challenge! To me, this is even bigger than our historic IMO-gold achievement last year; these problems challenge even top mathematicians. We share our results transparently, see paper and full thoughts in the thread. 👇 — https://nitter.net/lmthang/status/2026689272456294850#m
-
Model Failure: Debugging Your AI System’s Critical Logs
By
–
Your model is busted and the LOG is staring back at you… what’s YOUR play?