Introducing Devstral Small and Medium 2507! This latest update offers improved performance and cost efficiency, perfectly suited for coding agents and software engineering tasks.
OPEN SOURCE
-
Multi-Agent Architecture for Improved Model Performance
By
–
By the way, you can basically make the "Grok heavy" version of any model by having multiple agents running tools in parallel, then checking notes together and deciding which one is the best answer. I may release an open source project for that.
-
Hugging Face SmolLM3: Excellent 3B Model Release
By
–
And shoutout to friends at @huggingface for their SmolLM3 release! It's an excellent 3B model if you want to go bigger, and I loved reading their blog post (cc @eliebakouch and team)
-
Liquid Foundation Models V2 Generative AI Series Released
By
–
Here's the link to our blog post: https://
liquid.ai/blog/liquid-fo
undation-models-v2-our-second-series-of-generative-ai-models
… And you'll find below a link to the @huggingface models. Stay tuned, we have many more announcement to make! -
Local LLM Apps: Decentralized Alternative to Big Tech Monopolies
By
–
I'm really happy with what we've accomplished with LFM2. I'd love to see a future where everybody makes LLM-powered apps with efficient, local models that you own. It's so much sexier than a monopoly of giant companies stealing your data.
-

Fine-tuning tiny AI models with Colab notebooks tutorial
By
–
Note that these models are tiny, so I would always recommend fine-tuning them for your use case. My colleague @EdoardMosca created two Colab notebooks to help you in this adventure. You can find them in the model cards. More tutorials to come! 🙂
-

LFM2-1.2B outperforms larger Qwen3 model through efficient training
By
–
This speed should come at a cost but, thanks to efficient training, we managed to squeeze incredible performance in these tiny models. LFM2-1.2B is competitive with Qwen3-1.7B, a model 47% larger!
-

Liquid AI Open-Sources LFM2 Edge LLMs Generation
By
–
Liquid AI open-sources a new generation of edge LLMs! I'm so happy to contribute to the open-source community with this release on @huggingface
! LFM2 is a new architecture that combines best-in-class inference speed and quality into 350M, 700M, and 1.2B models. -
Huggingface explains stateless direct response MCP server choice
By
–
If you're developing MCP servers, you should give a read to how the @huggingface team built the Hub MCP, they explain why they chose a Stateless + Direct Response server over other options!
-
Hugging Face Launches Affordable Reachy Mini Desktop Robot
By
–
Hugging Face opened orders for Reachy Mini, an expressive, open-source desktop robot
— The Rundown AI (@TheRundownAI) 10 juillet 2025
Starting at just $299 and fully programmable in Python, it makes an ideal candidate for human-robot interaction, creative coding, and AI experimentationpic.twitter.com/C5AxiZdXyAHugging Face opened orders for Reachy Mini, an expressive, open-source desktop robot Starting at just $299 and fully programmable in Python, it makes an ideal candidate for human-robot interaction, creative coding, and AI experimentation
