Tackling FirstProof was our 7th math research paper, which was done autonomously and the solution to problem #7 is at Level 2 "Publishable Research" as well. nitter.net/lmthang/status/2026689… Thang Luong (@lmthang) Thrilled to share: #Aletheia, our math research agent, just solved 6/10 notoriously hard FirstProof problems autonomously, the best result in the inaugural challenge! To me, this is even bigger than our historic IMO-gold achievement last year; these problems challenge even top mathematicians. We share our results transparently, see paper and full thoughts in the thread. 👇 — https://nitter.net/lmthang/status/2026689272456294850#m
AI
-
Model Failure: Debugging Your AI System’s Critical Logs
By
–
Your model is busted and the LOG is staring back at you… what’s YOUR play?
-

Google Research Adds Benedek Rozemberczki to OMEGA Algorithms Team
By
–
Google Research has added Benedek Rozemberczki to its OMEGA Algorithms Research Team! #BigData #Analytics #AI #MachineLearning #DataScience #IoT #IIoT #Python #RStats #TensorFlow #JavaScript #ReactJS #CloudComputing #Serverless #DataScientist #Linux #Programming #Coding
-
Tesla Self-Driving Challenged by Unpredictable Portuguese Drivers
By
–
I am unsure Tesla self driving can handle Portuguese drivers, they drive like madmen and don't ever use blinkers
-
Revolutionary Setup Change for Better Coding Agents
By
–
Most people are using coding agents completely wrong. There's a simple setup change that makes them dramatically better. And once you switch to it, you'll never go back. Here's the full breakdown + the starter prompt to copy. Trust me, you NEED to try this. Matt Shumer (@mattshumer_) x.com/i/article/203505392109… — https://nitter.net/mattshumer_/status/2035058834117419036#m
→ View original post on X — @mattshumer_, 2026-03-20 18:21 UTC
-

REDSearcher: Novel Framework for Advanced LLM Search Agents
By
–
Tired of LLMs getting lost on deep, complex search tasks, and costing a fortune doing it? The REDSearcher Team presents a novel framework designed to make LLMs elite long-horizon search agents. It cleverly generates high-quality, complex search tasks, actively trains models to
-
High-end GPU recommendations for AI/ML workloads
By
–
while I agree with his general thinking about how we’ll rethink home design, seems optimistic to think that a meaningful percentage of homes would be redesigned in a 5-10 yr period https://t.co/OqXdOuhbRh
— Yohei (@yoheinakajima) 20 mars 2026while I agree with his general thinking about how we’ll rethink home design, seems optimistic to think that a meaningful percentage of homes would be redesigned in a 5-10 yr period
-

AI Agents: Unpredictable Production Behavior and New Testing Challenges
By
–
AI agents introduce a new production challenge: you simply don’t know what your agent will do until it’s in production. They interpret open-ended language, behave differently based on subtle shifts in phrasing, and take multi-step actions that are hard to predict in development
-
Aletheia AI Solves Problem Without Hints from HAI Card
By
–
You can take a look at the HAI card and the transcript. The author specified only the original problem, no hint was given to Aletheia.
-
Tesla Full Self-Driving Approved in Netherlands, EU Expansion Likely
By
–
Tesla self driving approved in Netherlands on April 10th After that every EU country can essentially "copy" the approval and instantly approve it in their country too So by end of 2026 it's likely we'll have Tesla Full Self Driving in most of the EU
