> Emu3: Next-token prediction conquers multimodal tasks This is the most important research in months: we’re now very close to having a single architecture to handle all modalities. The folks at BAAI just released Emu3, a single model that handles text, images, and videos all
LLMS
-

Manish Shah Presents SN40L LLM Solution at RIKEN AI Seminar
By
–
Our own Manish Shah spoke at the 5th Joint Seminar on AI Architectures at RIKEN in Tokyo, showcasing our cutting-edge SN40L solution for large language models. Thanks to Dr. Kentaro Sano and the @RIKEN_AIP_EN team for the insightful discussions! #LLM #AI #MachineLearning
-
OpenAI Releases Developer Tools with 98% Cost Reduction
By
–
shipping a few new tools for developers today! from last devday to this one: *98% decrease in cost per token from GPT-4 to 4o mini
*50x increase in token volume across our systems
*excellent model intelligence progress
*(and a little bit of drama along the way) -
Custom Code Assistants with Groq and CodeGPT
By
–
Tune in tomorrow to learn how to create custom code assistants with Groq and @codegptAI
. We'll explain how to use open source models as code assistants in your editor, create agents & enrich them with knowledge through RAG & Knowledge Graphs. -
Efficient logprob encoding using control tokens for low values
By
–
A more efficient way to do this could be to have typical logprobs encoded normally and exceptionally low ones marked with a control token as exact magnitudes don’t matter much beyond them being low — this is the approach of SequenceMatch
-

Mistral AI Releases Pixtral-12B Multimodal Model
By
–
Pixtral-12B: Mistral AI’s First Multimodal Model: Introduction Mistral has released its very first multimodal model, namely the Pixtral-12B-2409. This… https://
analyticsvidhya.com/blog/2024/09/p
ixtral-12b/?utm_source=dlvr.it&utm_medium=twitter
… #DataAnalytics #DataScience #DataDriven #IoT #MachineLearning #ITManager #SaaS #ArtificialIntelligence -
Add source highlighting to your RAG system for trust
By
–
> Add source highlighting to your RAG system! 📄💡
— m_ric (@AymericRoucher) 1 octobre 2024
RAG systems are supposed to make your LLM's answer more trustworthy, by inserting in the prompt some supporting documents from a knowledge base : we say that we're "adding some context".
👎 But if you don't know which part of… pic.twitter.com/5KmE7wcMua> Add source highlighting to your RAG system! RAG systems are supposed to make your LLM's answer more trustworthy, by inserting in the prompt some supporting documents from a knowledge base : we say that we're "adding some context". But if you don't know which part of
-
Pregen from base model with frequency for greetings
By
–
Pregen + prompt with a sample sounds like a good workaround to me; just pregen a lot so you don’t mind the diversity. You could even pregen from a base model and count frequency so common greetings are common.
-
LLama 3.2 1GB 2GB Versions Enable Ubiquitous Deployment
By
–
in case you missed it: LLama 3.2 in 1GB and 2GB sized version run it everywhere…
-
OpenAI Announces API Price Cuts and Advanced Features Updates
By
–
API Price Down
Custom GPTs API
Whisper Diarization
API With Search Ground(Search GPT?)
GPT-4o-Turbo
Advanced Voice Mode API
Prompt Caching
Workflow
