Agreed, MMLU Pro is not perfect either but we have enough data points to make a plot now
LLMS
-
PaliGemma Fine-tuning with AutoTrain Guide
By
–
blog: https://
huggingface.co/blog/abhishek/
paligemma-finetuning-autotrain
… github repo: -

AutoTrain Enables VLM Finetuning with PaliGemma Support
By
–
NEW TASK ALERT: VLM Finetuning AutoTrain just added VLM finetuning: Captioning and VQA for PaliGemma. Now, its super-easy to finetune PaliGemma on your own custom dataset. Which model and tasks would you like to see next? Let us know in Github Repo!
-

DeepSeek-V2 and Mistral Large 2 Added to Updated Figure
By
–
Due to popular demand, I've updated this figure to include DeepSeek-V2 and Mistral Large 2. It's also more zoomed for readability.
-
Llama 3.0 Update: Libraries Not Current, Long Context Impact
By
–
The libraries are not up-to-date yet, so it received the Llama 3.0 treatment. It's a good question, long context might have suffered depending on how it's handled, but I haven't tested it.
-

First Uncensored Llama 3.1 Instruct Model Released
By
–
I'm releasing the first uncensored Llama 3.1 Instruct model with GGUF quants This is an abliterated model that works well in my tests. You can learn more about it in my article. Model: https://
huggingface.co/mlabonne/Meta-
Llama-3.1-8B-Instruct-abliterated
… Article: https://
huggingface.co/blog/mlabonne/
abliteration
… -

Optimize RAG DSPy Applications for Better Performance
By
–
Optimize a RAG DSPy Application – Parea AI https://
bit.ly/3Lkd0xC
#AI #MachineLearning #DeepLearning #LLMs #DataScience -

Llama-3.1-8b Fine-Tuned Outperforms GPT-4 and Other Models
By
–
There's a new "best #SLM" in town! We fine-tuned Llama-3.1-8b-instruct on 25 tasks and it shows a huge improvement over #GPT-4, GPT-4o mini, fine-tuned #Phi-3, and fine-tuned #Mistral-7b. Small language models continue to set the standard for performance, cost, and privacy!
-
Mistral Large 2 Transformer Architecture
By
–
Mistral Large 2 is based on the Transformer architecture