With vLLM via pip, run any of the models in the Llama 4 family with a simple command.
OPEN SOURCE
-
Llama 4 Ecosystem: Rich Experiences and New Model Details
By
–
We can’t wait to see the rich experiences people build in the new Llama ecosystem! Even more details on the Llama 4 herd in the model card
-

Llama 4 supports 12 languages with fine-tuning options
By
–
Llama 4 supports 12 languages for tasks like multilingual writing — and developers can fine-tune Llama 4 models for additional languages beyond these 12, provided they comply with the Llama 4 Community License and the Acceptable Use Policy.
-

Llama 4 Scout and Maverick Powered by Llama 4 Behemoth Distillation
By
–
Llama 4 Scout and Llama 4 Maverick’s industry-leading performance is in large part thanks to distillation from Llama 4 Behemoth, our most powerful model yet. Be on the lookout for more details on Llama 4 Behemoth at a future date!
-

Llama 4: Mixture of Experts Architecture for Efficient Models
By
–
Llama 4 is our first collection of models built using a mixture of experts (MoE) architecture. This architecture is more compute efficient for model training and inference and delivers higher quality models compared to dense architectures.
-

Llama 4 Scout: State-of-the-art Performance with 10M Token Context
By
–
Llama 4 Scout delivers state-of-the-art performance for its class enabled by continued “mid-training” with new training recipes using specialized datasets enhancing model quality and unlocking a 10M token input context length.
-

Llama 4 Scout and Maverick Now Available on Poe
By
–
The first versions of Llama 4 Scout and Maverick are now available on Poe, thanks to @FireworksAI_HQ
!


