We've been busy in Q4, exhibiting and presenting at #SC23, #AppliedIntelligence, #EasternDefenseSummit, #ELC23, #GITEX2023, #STACSummit, #AISummit, and more! See the fastest way to run #LLMs, the #Groq #LPU ™ Inference Engine, at booth #1520 at #DoDIIS23 this week.
LLMS
-
BPE Tokenization Dies: 2024 Marks Major Transformer Evolution
By
–
i can see it now: 2024 will be remembered as the year BPE died tokenization is by far the clunkiest part of a transformer; one last remaining bit of inelegance in an otherwise hyperoptimized model architecture time for it to go
-
ChatGPT’s First Year: AI Evolution Since Launch
By
–
It's hard to believe, but ChatGPT is now over a year old. I tried to put a blog post about what has happened in the World of AI since then, and I quickly realized that a comprehensive overview would be beyond anyone's purview. Nonetheless, I managed to put together a piece on
-
ChatGPT Screenshot to Code GPT converts website screenshots to code
By
–
ChatGPT can be a powerful coding assistant..
— God of Prompt (@godofprompt) 11 décembre 2023
That's why I've built Screenshot to Code GPT.
Upload any screenshot of a website, and it will turn it to code in seconds!
Try it here: https://t.co/0dcCMJGR41 #CustomGPTs #GPTs #GPTPlus #GPTStore #ChatGPT4 #ChatGPTPlus #OpenAI… pic.twitter.com/ayqbSvpUV9ChatGPT can be a powerful coding assistant.. That's why I've built Screenshot to Code GPT. Upload any screenshot of a website, and it will turn it to code in seconds! Try it here: https://
godofprompt.ai/gpts/screensho
t-to-code-gpt
… #CustomGPTs #GPTs #GPTPlus #GPTStore #ChatGPT4 #ChatGPTPlus #OpenAI -

Mixtral Details and La plateforme Developer Platform Launch
By
–
More details about Mixtral can be found at https://
mistral.ai/news/mixtral-o
f-experts/
… We are also very happy to announce "La plateforme" our early developer platform (in beta & limited access), to access our models through our API: https://
mistral.ai/news/la-platef
orme/
… (7/n) -

Mixtral Outperforms Llama 2 70B on European Language Benchmarks
By
–
Mixtral has been trained on a lot of multilingual data and significantly outperforms Llama 2 70B on French, German, Spanish, and Italian benchmarks. (4/n)
-

Mixtral outperforms Mistral 7B in science, mathematics, and code
By
–
Compared to Mistral 7B, Mixtral is significantly stronger in science, in particular in mathematics and code generation. (5/n)
-
Mixtral Architecture: 8 Feedforward Blocks with Router-Selected Experts
By
–
Mixtral has a similar architecture as Mistral 7B, with the difference that each layer is composed of 8 feedforward blocks. For every token, at each layer, a router network selects two experts to process the current state and combine their outputs. (2/n)
-
Mixtral: 12B Speed with 45B Parameter Access via Expert Selection
By
–
Even though each token only sees two experts, the selected experts can be different at each timestep. As a result, Mixtral decodes at the speed of a 12B model, while effectively having access to 45B parameters. (3/n)
-

Mixtral 8x7B: Open Weight Mixture of Experts Model Released
By
–
Very excited to release our second model, Mixtral 8x7B, an open weight mixture of experts model.
Mixtral matches or outperforms Llama 2 70B and GPT3.5 on most benchmarks, and has the inference speed of a 12B dense model. It supports a context length of 32k tokens. (1/n)