For more information about Aletheia, powered by Gemini #DeepThink, and our works in AI for mathematical and scientific discovery at @GoogleDeepMind and @GoogleResearch, check out our announcement last week nitter.net/lmthang/status/2021631…! Thang Luong (@lmthang) 6 months in, after the IMO-gold achievement, I’m very excited to share another important milestone: AI can help accelerate knowledge discovery in mathematics, physics, and computer science! We’re sharing Two new papers from @GoogleDeepMind and @GoogleResearch that explore how Gemini #DeepThink together with agentic workflows can empower mathematicians and scientists to tackle professional research problems. Some highlights: The first paper built a research agent #Aletheia, powered by an advanced version of Gemini Deep Think, that can autonomously produce publishable math research and crack open Erdős problems. The second paper, built on similar agentic reasoning ideas, helped resolve bottlenecks in 18 research problems, across algorithms, ML and combinatorial optimization, information theory and economics. See the thread for details about the two papers and the joint blog post. — https://nitter.net/lmthang/status/2021631397614731563#m
AGENTS
-

Aletheia AI Solves Open Math Problem P7 Successfully
By
–
This is a remarkable milestone in which our agent can work on a research problem for a very long time, then come back and tell us if it has succeeded or failed! We visualize the inference cost Aletheia decided to spend on each candidate solution (as a multiple of the inference cost of for solving Erdős-1051, see our previous work nitter.net/lmthang/status/2018354…). P7 is extremely interesting. It has been an open problem for several years, and nobody else came close to solving it in the FirstProof contest per @tonylfeng. We initially thought Aletheia had no chance; turned out it was right! Aletheia spent most compute on P7, 16x amount we used for Erdős-1051. Remarkably, per @kimshmath, "This was the first case that I have ever seen that an AI applies several deep mathematical results (by Cartan/Leray/Borel/Atiyah/Quillen/Novikov/Kasparov…) flawlessly. It is a very unique instance."
-

Perplexity Computer: Unified AI System for End-to-End Project Management
By
–
Introducing Perplexity Computer.
— Perplexity (@perplexity_ai) 25 février 2026
Computer unifies every current AI capability into one system.
It can research, design, code, deploy, and manage any project end-to-end. pic.twitter.com/dZUybl6VkYIntroducing Perplexity Computer. Computer unifies every current AI capability into one system. It can research, design, code, deploy, and manage any project end-to-end.
→ View original post on X — @abhi1thakur, 2026-02-25 16:27 UTC
-

Aletheia Agent Solves 6 of 10 FirstProof Math Challenge Problems
By
–
Exciting results in AI math research! We use Aletheia agent, powered by Gemini 3 Deep Think, to tackle the FirstProof challenge. Operating completely autonomously, Aletheia successfully solved 6 out of the 10 problems. Check out the full paper for details on the methodology and expert evaluations. arxiv.org/abs/2602.21201
-

Aletheia Math Agent Solves Hard FirstProof Problems Autonomously
By
–
Thrilled to share: #Aletheia, our math research agent, just solved 6/10 notoriously hard FirstProof problems autonomously, the best result in the inaugural challenge! To me, this is even bigger than our historic IMO-gold achievement last year; these problems challenge even top mathematicians. We share our results transparently, see paper and full thoughts in the thread. 👇
-
Building an AI political agent team without calling them ministers
By
–
Read carefully. Building an AI political responsible agent is already possible and it works. Tomorrow, we will build an entire team of AI responsible political agents with skills and areas dedicated to each AI agent. We will not call them ministers,
-
Community manager jobs threatened by AI
By
–
Read carefully. Most jobs related to social media management, including the famous role of Community Manager that makes millions of under-35s dream, will disappear in the coming years in favor of fully autonomous AI agents.
-
LAP: Language-Action Pre-Training for Zero-shot Cross-Embodiment Transfer
By
–
LAP Language-Action Pre-Training Enables Zero-shot Cross-Embodiment Transfer
-
Reflective Test-Time Planning for Embodied LLMs
By
–
Learning from Trials and Errors Reflective Test-Time Planning for Embodied LLMs
-
Essential principles before launching your AI agent
By
–
Before you build and launch a single AI agent, finish this sentence:
— DataRobot (@DataRobot) 25 février 2026
"I shouldn't have to…" pic.twitter.com/QSjKIT76aoBefore you build and launch a single AI agent, finish this sentence: "I shouldn't have to…"
