The great thing is that the names are so baffling that the most important models OpenAI released were names davinci-002, GPT-3.5, GPT-4, o1-preview, o3, GPT-5 Pro, and you would never know the ways they are connected (my listing was chronological)
RESEARCH
-
Sakana AI Accelerates Sparse LLMs with NVIDIA
By
–
Sakana AIは、@NVIDIAとの共同研究で、スパースなTransformer言語モデルの推論・学習を高速化する新しいGPUカーネルとデータ形式を開発しました。
— Sakana AI (@SakanaAILabs) 9 mai 2026
ブログ:https://t.co/fMARMRFsJJ
LLMのコストの大部分を占めるフィードフォワード層では、実は各トークンに対して大半の活性がほぼゼロで無駄な計算に… https://t.co/nTMg0QgdSrSakana AI has developed new GPU kernels and data formats that accelerate inference and training of sparse Transformer language models through joint research with @NVIDIA
. Blog: https://
pub.sakana.ai/sparser-faster
-llms/
… In the feedforward layers, which account for the majority of LLM costs, most -

Claude Mythos Preview snapshot outperforms next best model by 2x
By
–

An early Claude Mythos Preview snapshot we provided METR has a time horizon of more than 2x the next best model on their 80% success rate benchmark
-

UFOs on HF: who will train first computer vision model?
By
–
The UFOs are on HF thanks to @MTSlive
! Who’s going to train the first computer vision model? https://
huggingface.co/MTSlive/datase
ts
… -
ChatGPT 5.5’s Math Skills Analyzed
By
–
I write about this in more detail in a blog post with a guest contribution from Isaac Rajagopal, a student at MIT on whose work ChatGPT built, who gives his assessment of the level of mathematical ability displayed by the model. https://
gowers.wordpress.com/2026/05/08/a-r
ecent-experience-with-chatgpt-5-5-pro/
… -

DeepMind solves 48% of a complex math benchmark
By
–
Google DeepMind's AI co-mathematician just scored 48% on FrontierMath Tier 4, a new high on a benchmark of 50 research-level math problems some professors expected AI wouldn't touch for decades. The system generated a proof so flawed its own reviewer flagged it as wrong. But
-
ByteDance AI Video Model Sparks China’s Next Big Moment
By
–
ByteDance’s new AI video model goes viral as China looks for second DeepSeek moment
#AI #AIio #AIInnovation #ML #DataScience #Futureofwork @timnitgebru @oriolvinyalsml @ceobillionaire @soumithchintala @waitin4agi_ @sallyeaves @bernardmarr -

DeepMind AI scores 48% on research-level math problems
By
–

DeepMind's AI co-mathematician scored 48% on FrontierMath Tier 4-research-level math problems that professional mathematicians need weeks to solve. The base model (Gemini 3.1 Pro) scores 19% alone. The entire jump comes from agentic scaffolding, parallel agents reviewing each
-

A Theory of Generalization in Deep Learning (arXiv link)
By
–
A Theory of Generalization in Deep Learning Elon Litman, Gabe Guo: https://
arxiv.org/abs/2605.01172 #ArtificialIntelligence #AIAgents #DeepLearning -

A Theory of Generalization in Deep Learning
By
–
A Theory of Generalization in Deep Learning Elon Litman, Gabe Guo: https://
arxiv.org/abs/2605.01172 #ArtificialIntelligence #AIAgents #DeepLearning