Seriously?! It seems like a new version of DeepSeek is about to drop soon. A V3, not R2. According to leaks, it could be on par with GPT-4.5—this would make it the best model available! I can't wait to see. I'll keep you updated as soon as I get more info.
LLMS
-

Deterministic Algorithms for LLM Sampling Beyond Stochasticity
By
–
The LLM parrots need not be stochastic. Long weekend thought motivated by a posting by @balazskegl : A few years ago @yutianc and @wellingmax introduced me to herding. It is a dynamical systems (deterministic) algorithm for producing samples for a multinomial discrete
-

LLMs and Humans Trade Compression for Meaning Analysis
By
–
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
Paper: https://
arxiv.org/pdf/2505.17117
v1.pdf
… -

How LLMs Organize Concepts Differently from Human Brains
By
–
Do LLMs Think Like Humans? New Study Says Not Quite This paper dives into a fascinating question:
Do Large Language Models organize concepts the way humans do?
Spoiler: They don’t—at least not yet. Humans use semantic compression—they categorize by balancing expressive -

State of the Art Evaluation in Code Generation
By
–
A pretty accurate eval of the state of the art in code generation I think we’ve all been there
-
Is Grok 3.5 Finally Around the Corner?
By
–
Bon, est-ce que Grok 3.5 serait enfin dans les parages ? Après tant d'annonces qui n'ont pas aboutit, plus rien n'est sur.
-
Subscribe to Weekly AI Engineering Tutorials on ML and LLMs
By
–
If you're interested in ML, LLMs, and AI Agents and want to receive tutorials every week, subscribe to AI Engineering (for free):
-
Beyond LLMs: AI’s Role in Scientific Discovery
By
–
In a recent interview I talk about what it takes for AI to make new scientific discoveries. tldr: it won’t be just LLMs.
-

Comparative Bias Analysis Datasets for Popular Chatbots
By
–
Would be useful to have similar datasets for GPT, Claude, Grok, and most of the popular chatbots to better analyze how biased they are
