ChatGPT isn’t even 3 years old and going straight for Apple’s jugular. The trajectory of OpenAI continues to be completely unprecedented.
LLMS
-
Mistral Releases Devstral: Pareto-Optimal Open-Source Coding Model
By
–
https://
mistral.ai/news/devstral we’ve released an open-source model which is great at agentic coding tasks – by far Pareto optimal and a good prelude to what’s coming next -
Llama Models by Hand Workshop with Prof Tom Yeh
By
–
One of our most popular paper clubs in a long time: https://
youtu.be/VH1tSGIe5e0 Llama 1/2/3/4 by Hand with @ProfTomYeh is now live! -
Training Data Overlap Between BERT and T5 Models
By
–
yeah, this is a great question – like even though BERT and T5 are different for example they probably have a large amount of training data overlap
-
Google Gemini 2.5 Pro Establishes Google as AI Leader
By
–
The G letter in Google in GOAT.
Google Gemini 2.5 Pro is all you need. Finally shown all the other companies who is the leader in the AI space. -
Gemma Models Progress: Medical, ASL, and Dolphin Applications
By
–
Some great progress on open versions of Gemma models, whether you care about medical applications, American Sign Language, or dolphin.
-
Gemma Multimodal Models Progress on American Sign Language Translation
By
–
Here's one good sign: there's been some really nice progress on adapting our open Gemma multimodal models to translate American Sign Language to English. https://t.co/FBUvHoKMfO
— Jeff Dean (@JeffDean) 21 mai 2025Here's one good sign: there's been some really nice progress on adapting our open Gemma multimodal models to translate American Sign Language to English.
-

Strong Platonic Representation Hypothesis: Cross-Model Translation
By
–
theoretically, the implications of this seem big. we call it The Strong Platonic Representation Hypothesis: models of a certain scale learn representations that are so similar that we can learn to translate between them, using *no* paired data (just our version of CycleGAN)
-
Scale Convergence: Different AI Models Learn Identical Representations
By
–
a lot of past research (relative representations, The Platonic Representation Hypothesis, comparison metrics like CCA, SVCCA, …) has asserted that once they reach a certain scale, different models learn the same thing this has been shown using various metrics of comparison
-
Reinforcement Fine-Tuning LLMs with GRPO Short Course
By
–
New Course: Reinforcement Fine-Tuning LLMs with GRPO!
— Andrew Ng (@AndrewYNg) 21 mai 2025
Learn to use reinforcement learning to improve your LLM performance in this short course, built in collaboration with @Predibase, and taught by @TravisAddair, its Co-Founder and CTO, and @grg_arnav, its Senior Engineer and… pic.twitter.com/j5AXn3swADNew Course: Reinforcement Fine-Tuning LLMs with GRPO! Learn to use reinforcement learning to improve your LLM performance in this short course, built in collaboration with @Predibase
, and taught by @TravisAddair
, its Co-Founder and CTO, and @grg_arnav
, its Senior Engineer and