Cox also gave a special shoutout to our partners, who’ve helped make it possible to get started building with Llama right off the bat—regardless of your development environment. Because the Llama stack is modular, it’s easy to plug and play and switch between component parts.
LLMS
-
Llama 4 Delivers High Performance in Compact Package for Business Scaling
By
–
And we want to help businesses scale. That’s why Llama 4 packs a lot of performance into as small a package as possible.
-

Llama Reaches 1.2 Billion Downloads Milestone
By
–
We announced Llama’s first 1 billion downloads last month, and we’re updating that number to 1.2 billion downloads today. “And what’s cool if you look at Hugging Face where the downloads are happening is that most of these are actually Llama derivatives. We have thousands of
-
Meta’s Open Source Legacy Highlighted at LlamaCon 2025
By
–
-

R-Sparse: Rank-Aware Activation Sparsity for LLM Inference
By
–
R-Sparse: Rank-Aware Activation Sparsity for Efficient LLM Inference
Paper: https://
arxiv.org/pdf/2504.19449
v1
…
Code: https://
github.com/VITA-Group/R-S
parse
… -

R-Sparse: 50% faster LLM inference without retraining
By
–
What if your LLM ran twice as fast—with zero retraining?
R-Sparse slashes compute by 50% using a clever trick: skip the unimportant math without guessing what to skip. No tuning, no ReLU, no problem. It’s efficient inference, reimagined for the edge. -
Qwen3 Hybrid Reasoning System Enables Flexible Problem-Solving Approaches
By
–
Qwen3's hybrid reasoning system lets you switch between quick answers and thorough problem-solving. Use –no_think for rapid responses or –think for detailed reasoning (enabled by default). (2/3)
-
Qwen3-235B-A22B Now Available on Poe Platform
By
–
You can try Qwen3-235B-A22B at https://
poe.com/Qwen3-235B-A22
B-FW
… and across all platforms. (3/3) -

Alibaba launches Qwen3 powerful reasoning model on Poe
By
–
Now on Poe: Qwen3! This powerful new model from Alibaba combines reasoning and performance to excel at complex tasks. (1/3)
-
LlamaCon Event: Groq Accelerates Meta’s Llama API
By
–
LlamaCon in full swing! Swag, friends, customer demos from @MaitaiAI
, the @AIatMeta Llama API accelerated by Groq. It's a good day
