Super excited to have @b_roziere join Mistral to lead our code generation team and build the next generation of Codestral models!
@guillaumelample
-

Mistral Large 2 Excels in Coding, Math, and Hard Prompts
By
–
Mistral Large 2 (2407) is now on @lmsysorg
. It performs extremely well in the Coding, Hard Prompts, Math, and Longer Query categories, where it outperforms GPT4-Turbo and Claude 3 Opus. It is also doing very well in Instruction Following where it ranks above Llama 3.1 405B. -

Mistral Large 2 Now Available Free on Le Chat
By
–
You can use Mistral Large 2 on Le Chat — it's free! https://
chat.mistral.ai -
Mistral Large Instruct 2407 Model Released on HuggingFace
By
–
The model is available (for research purposes only!) on HuggingFace: https://
huggingface.co/mistralai/Mist
ral-Large-Instruct-2407
…
Blog post: -

Mistral Large Improved Alignment and Instruction Capabilities Performance
By
–
Compared to the previous Mistral Large, much more effort was dedicated to alignment and instruction capabilities. On WildBench, ArenaHard, and MT Bench, it performs on par with the best models, while being significantly less verbose. (4/N)
-

Mistral Large 2 outperforms Llama 3.1 on Multilingual MMLU
By
–
On Multilingual MMLU, the performance of Mistral Large 2 significantly outperforms Llama 3.1 70B base (+6.3% average over 9 languages) and is on par with Llama 3 405B (-0.4% below). (3/N)
-

Mistral Large 2 outperforms Llama 3.1 405B on coding benchmarks
By
–
On HumanEval and on MultiPL-E, Mistral Large 2 outperforms Llama 3.1 405B instruct, and scores just below GPT-4o. On MATH (0-shot, without CoT) it only falls behind GPT-4o.
(2/N) -
Mistral Large 2: New 123B Model Outperforms Llama 3.1 405B
By
–
Today, we release Mistral Large 2, the new version of our largest model. Mistral Large 2 is a 123B-parameter model with a 128k context window. On many benchmarks (notably in code generation and math), it is superior or on par with Llama 3.1 405B. Like Mistral NeMo, it was trained
-
Mistral Nemo Model Now Available on Hugging Face
By
–
The model is also available on Hugging Face:
Base model: https://
huggingface.co/mistralai/Mist
ral-Nemo-Base-2407
…
Instruct model: -

Mistral NeMo 12B Model Release with NVIDIA Collaboration
By
–
Very happy to release our new small model, Mistral NeMo, a 12B model trained in collaboration with @nvidia
. Mistral NeMo supports a context window of 128k tokens, comes with a FP8 aligned checkpoint, and performs extremely well on all benchmarks. Check it out!