Less than 1 week until the 7th workshop on Neural Scaling Laws at ICML 2024 Tatiana Shavrina, PhD, Research Scientist Manager at Meta, will present on the latest advancements and limitations in multilingual language models, highlighting the significant milestones in machine
GENERATIVE AI
-

Small Language Models: Efficient Alternatives to Large Language Models
By
–
Small language models (SLMs) are often overlooked in favor of their larger counterparts (LLMs) like @ChatGPTapp
. But did you know SLMs can be more efficient, cost-effective, and perfect for specialized tasks? Have you considered using SLMs for your business needs? #AI #SLM -

Dissecting LLM failures to understand their causes
By
–
I don’t post LLM failures (like the one below) because I think LLMs are bad or even overhyped; I do it because: 1) I think dissecting concrete LLM failures to understand their causes is a tragically underused approach to fixing them 2) previously well known issues in LLMs (e.g.
-

Mistral Releases Mathstral 7B for Scientific Problem Solving
By
–
NUEVO MODELO MATHSTRAL 7B! Un Mistral entrenado para la mejor resolución de problemas científicos y matemáticos. Disponible para descargar y jugar con él 🙂
-

LLM Fine-tuning vs Contextualization: When to Choose
By
–
Encadrez juste ça #LLM #Context #matrix ça vous dira si vous avez besoin de fine tuner ou de mieux contextualiser
-
AI Capabilities and Limitations in Complex Problem Solving
By
–
Good point Tom. It feels more like the combinatorial type, or just as a way to get creators started. This is widely reported and I’ve experienced it myself. I’m still waiting for a model that will design a fusion reactor that solves our energy problems, but such AI is still
-
Llama 3 Model Size Increase: 7B to 8B Explanation
By
–
Some context on why our smallest Llama 3 model went from 7B → 8B. More details on the changes to the tokenizer in the full conversation with @astongzhangAZ ➡️ https://t.co/cKuUfwmuZ4 pic.twitter.com/OfqCJtfSJW
— AI at Meta (@AIatMeta) 16 juillet 2024Some context on why our smallest Llama 3 model went from 7B → 8B. More details on the changes to the tokenizer in the full conversation with @astongzhangAZ https://
youtu.be/Tmdk_H2WDj4 -

Claude Android App Now Available on Google Play
By
–
The Claude Android app is now available.
— Anthropic (@AnthropicAI) 16 juillet 2024
Download on Google Play: https://t.co/tRJJ1xDScn pic.twitter.com/ZnKqQJJUwKThe Claude Android app is now available. Download on Google Play: http://
anthropic.com/android -
AI Empowers Creativity: 84% of Users Report Enhanced Creative Capabilities
By
–
“84% of AI users said it made them more creative”. AI is tool to bolster creativity. I believe our kids will run with this idea to create wonderful new things that we can’t even imagine today.
-

Mistral Releases Mathstral 7B and Codestral Mamba Models
By
–
Today we are releasing two small models: Mathstral 7B and Codestral Mamba 7B. On the MATH benchmark, Mathstral 7B obtains 56.6% pass@1, outperforming Minerva 540B by more than 20%. Mathstral scores 68.4% on MATH with majority voting@64, and 74.6% using a reward model. Codestral
