GPT 5.4 is cheaper and more performant for xls and deep research than other SOTA models We have now managed to improve performance by 33% on other top Gemini or Claude models OpenAI needs to double down on this win and launch GPT 6.0 ASAP
LLMS
-
Uni-1 model integrates reasoning and image generation in a continuous process
By
–
La IA ya no necesita dos pasos para crear.
— SONIA (@S0N_IA_) 23 mars 2026
Uni-1 unifica el pensamiento y la generación de píxeles en un único proceso continuo. Como imaginar y dibujar al mismo tiempo.
Esto es lo que viene después de los LLMs. https://t.co/ljsRMHMoCWLa IA ya no necesita dos pasos para crear. Uni-1 unifica el pensamiento y la generación de píxeles en un único proceso continuo. Como imaginar y dibujar al mismo tiempo. Esto es lo que viene después de los LLMs.
-

GradMem: Gradient Descent Memory for Context Compression
By
–
"GradMem: Learning to Write Context into Memory with Test-Time Gradient Descent" Instead of giving LLMs huge KV caches or one-shot compressed summaries, GradMem shows that you can let a frozen model take some test-time gradient steps to write a long context into a small memory.
-

GPT-5.4 Pro Solves FrontierMath Open Problem
By
–
That is really really impressive: GPT-5.4 pro has solved one of the open problems in FrontierMath. Kevin Barreto and Liam Price, using GPT-5.4 Pro, produced a construction that Will Brian confirmed, with a write-up planned for publication We are accelerating
-

Large-Scale Online Deanonymization Using LLMs Research
By
–
Large-scale online deanonymization with LLMs Lermen et al.: https://
arxiv.org/abs/2602.16800 #ArtificialIntelligence #AIAgents -

Large-scale deanonymization attacks using language models
By
–
Large-scale online deanonymization with LLMs Lermen et al.: https://
arxiv.org/abs/2602.16800 #ArtificialIntelligence #AIAgents -

Plan and Budget Framework Optimizes LLM Token Efficiency
By
–
Is your LLM wasting valuable tokens "overthinking" or "underthinking" complex tasks? MIT CSAIL, Virginia Tech, University of Virginia, and Michigan State University present Plan and Budget. This new framework helps LLMs decompose complex problems into manageable sub-questions
-
Uni-1 understands the real and eliminates the need for perfect prompts
By
–
Uni-1 → understands how something real should look This changes something key: You don't need to perfect prompts.
-

AI Solves First Open Problem from FrontierMath Benchmark
By
–
EPIC! It's confirmed that the AI has solved the first mathematical problem from the FrontierMath: Open Problems benchmark, which consists of problems that remained unsolved despite attempts by the mathematical community. The first of many more to come!
-
Optimizing Default LLM Models for OpenClaw AI Tool Usage
By
–
For those using OpenClaw at a high level, what’s your favorite default model? I was using Nemotron 3-super locally on my Spark but it hits context limits too quickly. I’m mostly using Sonnet-4.6 now but API costs rack up fast. I love my claw but honestly haven’t optimized models