seems like we managed to make a blog-post about data processing go viral faith in humanity restored
@thom_wolf
-

Understanding High Performance Language Models: Llama3 GPT-4 Mixtral
By
–
This might be one of the most important 45-mn read you could indulge in today if you want to understand the secret behind high performance large language models like Llama3, GPT-4 or Mixtral Inspired by the @distillpub interactive graphics papers, we settled to write the most
-
LHC Black Holes to AGI: History Repeating Fear Cycles
By
–
this same LHC which was feared to potentially create black holes that would destroy the earth in mainstream news 🙂 –history keep repeating itself. Yesterday fearing black-holes from uncontrollable particules-collider, today fearing terminator-like risk from uncontrollable AGI
-
FineWeb Extended Report and Blog Post Coming Soon
By
–
We're finishing a pretty awesome extended blogpost/report on FineWeb in which you'll find all the info – stay tuned!
-

Golden Gate Claude Version Offers New Interpretability Features
By
–
If you enjoyed the interpretability blog post I shared yesterday, go play with the temporary available “Golden Gate” version of Claude – you’ll be surprised https://
x.com/elytramithra/s
/elytramithra/status/1793916830987550772
… -

Deep Issue in AI Field: GenAI Image Quality Problem
By
–
great (short) read from Sasha documenting this deep issue in our AI field no need to go far to find examples btw, just scroll here on X and look at GenAI pictures for bots and ads
-

Anthropic Interpretability Paper on Scaling Monosemanticity
By
–
The new interpretability paper from Anthropic is totally based. Feels like analyzing an alien life form. If you only read one 90-min-read paper today, it has to be this one https://
transformer-circuits.pub/2024/scaling-m
onosemanticity/index.html
… -

Microsoft Releases Phi-3 Series: 3.8B to 14B Models
By
–
A new series of Phi-3 mini, small & medium models (MIT license (3.8B – 7B – 14B) Long-context
Phi-3-mini-128k-instruct: https://
huggingface.co/microsoft/Phi-
3-mini-128k-instruct
…
Phi-3 small 128k: https://
huggingface.co/microsoft/Phi-
3-small-128k-instruct
…
Phi-3 medium 128k: https://
huggingface.co/microsoft/Phi-
3-medium-128k-instruct
… Multimodal:
Phi-3-vision-128k-instruct: -
Open AI as Common Good: Ownership and Transparency Over Closed Systems
By
–
like imagining a future where AI is something people can own, understand, tweak, build and improve together instead of a closed/black-box product full of secrets A common-good knowledge built on top of open-standard like the internet is or even science it self (no one privately
-
Moving beyond lonely futures: reimagining AI product inspiration
By
–
oh no but also, what about finding inspiration for AI products in other places than movies picturing futures full of loneliness? (As much as I enjoyed the aesthetic of the movie it’s still a sad tale)