Starting today, open source is leading the way. Introducing Llama 3.1: Our most capable models yet. Today we’re releasing a collection of new Llama 3.1 models including our long awaited 405B. These models deliver improved reasoning capabilities, a larger 128K token context
GENERATIVE AI
-
Llama 3.1 405B: Training the Largest Model at Scale
By
–
Training a model as large and capable as Llama 3.1 405B was no simple task. The model was trained on over 15 trillion tokens over the course of several months requiring over 16K @NVIDIA H100 GPUs — making it the first Llama model ever trained at this scale. We also used the 405B
-
Fine-tuning Llama-3.1-8B with High-Quality Reasoning Datasets
By
–
What I'm excited about: Use Llama-3.1-405B to generate a high quality reasoning dataset, then fine-tune Llama-3.1-8B models.
— Jiquan Ngiam (@JiquanNgiam) 23 juillet 2024
Imagine having 4o-mini performance models, but open sourced, and tuned for your use cases. pic.twitter.com/mf9tMZ8vWMWhat I'm excited about: Use Llama-3.1-405B to generate a high quality reasoning dataset, then fine-tune Llama-3.1-8B models. Imagine having 4o-mini performance models, but open sourced, and tuned for your use cases.
-
Meta Releases New Llama Model Update Blog Post
By
–
Os dejo el link con el blog post de Meta. Yo voy a repasar la info y más que hacer un hilo, luego abrimos directo y comentamos la importancia de esta noticia 🙂 https://
llama.meta.com -
Llama 3.1 Now Available in Three Sizes: 8B, 70B, 405B
By
–
¡LLAMA 3.1 YA ESTÁ AQUÍ! En sus tres tamaños… 8B, 70B y 405B
-

AI Breakthroughs: Supercomputers, Weather Models, Health Tech
By
–
Top stories in AI today: -The "world's most powerful" supercomputer
-Google’s AI-powered weather model
-Transform images in seconds with Freepik Expand
-MIT's AI identifies breast cancer risk
-6 new AI tools & 4 new AI jobs Read more: http://
therundown.ai/p/elon-reveals
-grok-3
… -

VideoPoet Wins ICML Best Paper Award for Zero-Shot Video
By
–
Congratulations to the authors of "VideoPoet: A Large Language Model for Zero-Shot Generation" for winning one of this year's @icmlconf Best Paper Awards! #ICML2024 Paper: https://
openreview.net/forum?id=LRkJw
PIDuE
… Blog post: https://
goo.gle/4atanoj -

Distilling Billion-Parameter Models Into Efficient Online Versions
By
–
@smritichirps
, Andrew Gilchrist-Scott & Brad Stocks will be back at the #ICML2024 Google Research booth at 4pm CEST to explain how we distill billion+ parameter models into quick and resource-efficient online versions while maintaining their underlying world & language knowledge! -

Distilling Billion-Parameter Models into Efficient Online Versions
By
–
@smritichirps
, Andrew Gilchrist-Scott & Brad Stocks will be back at the #ICML2024 Google Research booth at 4pm CEST to explain how we distill billion+ parameter models into quick and resource-efficient online versions while maintaining their underlying world & language knowledge! -

Google Translate Adds 110 Languages Reaching 614 Million Speakers
By
–
Google Translate has announced its biggest update yet, adding 110 new languages to the platform. The new languages represent over 614 million speakers, making translations accessible to around 8% of the world's population. What do you think about Google Translate's massive
