this exceptional investigation into generative AI bias by @Leonardonclt and @dinabass is free to read (for a short time). please check it out! it's fantastically well done from tip to toe.
AI
-
Meta Scales Speech Technology to 1,100+ Languages Globally
By
–
Our work on the Massively Multilingual Speech (MMS) project has scaled speech-to-text & text-to-speech to support 1,100+ languages + trained new language identification models that can identify 4,000+ languages!
— AI at Meta (@AIatMeta) 12 juin 2023
More info & access to pretrained models ⬇️Our work on the Massively Multilingual Speech (MMS) project has scaled speech-to-text & text-to-speech to support 1,100+ languages + trained new language identification models that can identify 4,000+ languages! More info & access to pretrained models
-

Oracle Earnings: Key Takeaway on AI Strategy
By
–
Doing the whole "this is really all you need care about" thing again for the Oracle earnings this time today with this line right here:
-

Scale AI and Hugging Face advance LLM evaluation SOTA
By
–
We @scale_AI collaborated with @huggingface to drive forward the SOTA on LLM and AI evaluation. See @natolambert
's excellent Twitter thread to learn more! -
AI Generative Fill Extends Famous Animal Memes by parh0
By
–
AI generative fill extending famous animal memes by parh0
— AK (@_akhaliq) 12 juin 2023
instagram: https://t.co/7dXgW1MK1z pic.twitter.com/8ZzMovz6sUAI generative fill extending famous animal memes by parh0 instagram: https://
instagram.com/p/CtY6smpId81/ -
SlimPajama Dataset: Preprocessing Library for LLM Training
By
–
SlimPajama dataset – https://
lnkd.in/gCchZ-xz
Preprocessing library: https://
lnkd.in/gV7r3YNC
Read our blog: -

SlimPajama: Open-Source Cleaned RedPajama Dataset Released
By
–
We recently announced the availablity of SlimPajama – an open-source, cleaned, and deduplicated version of RedPajama-1T. It is half the size and trains twice as fast and when upsampled, performs equal or better than RedPajama. See below for the dataset and preprocessing library
-
The Reading Paradox: Want Books but Skip Reading
By
–
Everybody wants to write a book. Everybody wants to say they’ve read a book. Nobody actually wants to read a book.
-

Technological Advances Behind Large Language Models Take Years
By
–
Technological advances don’t happen overnight; the technology that helped make it possible for us to have large language models today (like that which underpins ChatGPT) came out six years ago. https://t.co/Z12ZglKNPN
— Rachel Metz (@rachelmetz) 12 juin 2023Technological advances don’t happen overnight; the technology that helped make it possible for us to have large language models today (like that which underpins ChatGPT) came out six years ago.
-

GPT4All: Free Local LLM for Document Analysis
By
–
GPT4All is an ecosystem to train and deploy powerful and customized large language models that run locally on consumer grade CPUs. GPT4All is the Local ChatGPT for your documents… and it is free! https://
bit.ly/45U1h1P