GPT-5.5 delivers this step up in intelligence without compromising on speed. GPT-5.5 matches GPT-5.4 per-token latency in real-world serving, while performing better across nearly every evaluation we measured. It also uses significantly fewer tokens to complete the same Codex
MACHINE LEARNING
-
GPT-5.5 Advances in Coding and Knowledge Work
By
–
GPT-5.5 excels at writing and debugging code, researching online, analyzing data, creating documents and spreadsheets, operating software, and moving across tools until a task is finished. The gains are especially clear in agentic coding, computer use, knowledge work, and early
-
Complex Investigation Reveals Root Causes and System Confounders
By
–
We take these reports incredibly seriously. In my time on the team, this has probably been the most complex investigation we’ve had. The root causes were not obvious, and there were many confounders.
-
GPT-5.5 Spud Launches: Major Leap in Coding and AI
By
–
BREAKING:
— Dan Shipper 📧 (@danshipper) 23 avril 2026
GPT-5.5 "Spud" is out and it is a BEAST
We've been testing it @every for the last 3 weeks on everything from coding, to writing, to knowledge work. Here's our day 0 vibe check:
– It's a step change in coding AND it's easy to talk to. It's fast and friendly and… pic.twitter.com/lMl2wdwADbBREAKING: GPT-5.5 "Spud" is out and it is a BEAST We've been testing it @every for the last 3 weeks on everything from coding, to writing, to knowledge work. Here's our day 0 vibe check: – It's a step change in coding AND it's easy to talk to. It's fast and friendly and
-
AI Model Could Speed Rare Disease Diagnosis
By
–
New Artificial Intelligence Model Could Speed Rare Disease Diagnosis
#AI #AIio #AIInnovation #ML #DataScience #Futureofwork @HaroldSinnott @fogoros @iainljbrown @NandoDF @katecrawford @drhassanrashidi @YuHelenYu -

DR-Venus: Frontier Edge-Scale Deep Research Agents with 10K Data
By
–
DR-Venus Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data paper: https://
huggingface.co/papers/2604.19
859
… -

Near-Future Policy Optimization Research Paper
By
–
Near-Future Policy Optimization paper: https://
huggingface.co/papers/2604.20
733
… -

RLVR Effectiveness in Low Data Compute Regimes Study
By
–
Our MLSys 2026 paper is live on arXiv: “Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes.” @realjustinbauer @Walshe_tech @pham_derek @harit_v @ArminPCM @fredsala and @paroma_varma present a comprehensive empirical study of open-source SLMs
-
AI Discovery of MRSA-Effective Pill Could Transform Treatment
By
–
If AI can find a pill effective vs MRSA that would be big.
-

MIT Improves Reasoning Model Confidence Calibration Through RL Training
By
–
How do top reasoning models become overconfident? MIT found that RL rewards correct answers w/o considering how sure the model is. By training them to estimate their confidence about each answer, the team boosted uncertainty estimates w/o hurting accuracy:
