All models are available today on @huggingface
! If you have a precise skill in mind for a task-specific model, please let me know in the comments!
LLMS
-
All AI models now available on Hugging Face platform
By
–
-
Model Release Enables Developers to Build AI Apps Locally
By
–
I'm happy it's released, it was such a cool post-training project Now I hope that app developers can take these models and build cool stuff with them, with less reliance on costly and slow cloud models. There's so much to do, life is too short to optimize ChatGPT prompts!
-

350M Math Reasoning Model Outperforms Larger Competitors
By
–
Finally, another personal favorite: a 350M math reasoning model. It turned out to be an absolute beast, outperforming Qwen3-0.6B and MobileLLM-R1. Note that it's also trained to output short CoTs, as a continuation of our work in this blog post: https://
liquid.ai/research/lfm-1
b-math-can-small-models-be-concise-reasoners
… -

LFM2-1.2B Achieves Parity with Larger Models for Edge
By
–
For latency purposes on edge devices, we wanted it to be a non-thinking model. That was a big challenge, but we managed to squeeze a ton of performance from LFM2-1.2B and perform on par with much bigger models on our internal bench (see figure) but also BFCL v3 and v4.
-

RAG Model Excels at Long Context Processing and Information Retrieval
By
–
We also have a RAG model It's particularly good for processing long contexts and retrieving relevant information. It's very flexible and can adapt to various input styles.
-

350M Model Shows Power of Fine-Tuning for Big Data
By
–
It comes in two sizes: 1.2B and 350M I'm a big fan of the 350M model that can be used on GPUs to do big data operations. It's an absolute banger that shows how powerful fine-tuning can be
-

Tiny Task-Specific AI Models for Edge Devices
By
–
We're releasing a collection of tiny task-specific models Want to do data extraction, translation, RAG, tool use, or math on a Raspberry Pi? We got you covered! Here are a few examples ↓
-
Using XML to improve Claude model performance
By
–
XML is a great way to get best results from Claude
-
100+ Free Step-by-Step AI Agents and RAG Tutorials
By
–
100+ free step-by-step tutorials with code covering: AI Agents RAG Systems Voice AI Agents MCP AI Agents Multi-agent Teams Autonomous Game Playing Agents P.S: Don't forget to subscribe for FREE to access future tutorials.
-
Nemotron: Building Open Source AI for GPU Systems
By
–
“There's two reasons why we're building Nemotron. The first is because it helps us build the GPUs and systems for the future.”
— NVIDIA (@nvidia) 25 septembre 2025
The second reason is because trust is the foundation of impactful AI–and it starts with understanding. That’s why Nemotron was built open source at its… pic.twitter.com/H1ih4x1WBi“There's two reasons why we're building Nemotron. The first is because it helps us build the GPUs and systems for the future.” The second reason is because trust is the foundation of impactful AI–and it starts with understanding. That’s why Nemotron was built open source at its