Lastly, La Platforme got a free tier for developers to experiment with Mistral APIs as well as an improved Mistral Small Model Source: https://
mistral.ai/news/september
-24-release/
…
LLMS
-

Mistral AI Releases Improved Small Model and Free API Tier
By
–
-

Mistral AI Releases Pixtral Vision Model on Le Chat
By
–

BREAKING: Mistral AI made Pixtral model (with vision capabilities) available on Le Chat along with other platform updates
-
Co-LLM Collaboration Improves Mathematical Problem Solving Accuracy
By
–
The base LLM tried to solve math problems like “a^3 · a^2 if a=5,” where it incorrectly calculated the answer to be 125. Co-LLM trained the model to collaborate w/the large math LLM Llemma, and together they determined that the correct solution was 3,125.
-
Co-LLM Outperforms Fine-Tuned Models in Collaborative AI
By
–
Co-LLM gave more accurate replies than fine-tuned simple LLMs & untuned specialized models working independently. Co-LLM can guide two models that were trained differently to work together, whereas other effective LLM collaboration approaches, such as "Proxy Tuning," need all of
-
MIT Algorithm Trains Local LLMs for Enterprise Document Management
By
–
Eventually, the MIT algorithm could assist w/enterprise documents, using the latest info it has to update them accordingly. Co-LLM could also train the small LLM on documents that must remain within the server, and then we can use both models for many different cases.
-
Co-LLM: Collaborative Expert Switch for Enhanced Token Generation
By
–
If you asked Co-LLM to name some examples of extinct bear species, two models would draft answers together. The general-purpose LLM begins to put together a reply, w/the switch variable intervening where it can slot in a better token from the expert mode (i.e. adding the year
-

Co-LLM Combines Base and Expert Models for Biomedical Applications
By
–
To showcase Co-LLM’s flexibility, the researchers used data like the BioASQ medical set to couple a base LLM w/expert LLMs in different domains, like the Meditron model, which is pre-trained on unlabeled medical data. This enabled Co-LLM to help answer inquiries a biomedical
-
Co-LLM Switch Variable Routes Tasks Between Base Expert Models
By
–
To decide when a base model needs help from an expert model, Co-LLM uses machine learning to train a "switch variable," or a tool that can indicate the competence of each word w/i the two LLMs’ responses.
— MIT CSAIL (@MIT_CSAIL) 17 septembre 2024
It’s like a project manager deciding when to call in a specialist. pic.twitter.com/E2UMzf9b4oTo decide when a base model needs help from an expert model, Co-LLM uses machine learning to train a "switch variable," or a tool that can indicate the competence of each word w/i the two LLMs’ responses. It’s like a project manager deciding when to call in a specialist.
-

MIT Co-LLM Algorithm Enables Specialized Model Collaboration
By
–
Can LLMs learn to "phone a friend?" MIT CSAIL’s new "Co-LLM" algorithm can pair a general-purpose base LLM w/a more specialized model & help them work together. It reviews each token & sees where it needs to call upon an expert, leading to more accurate & efficient replies to
-

Llama Model Shows Strong Growth and Increasing Cloud Adoption
By
–
We recently shared an update on the growth of the Llama. Tl;dr: downloads are growing fast, our major cloud partners are seeing rapidly increasing usage of Llama on their platforms and we're seeing great adoption across industries! Read the full update https://
go.fb.me/d01004