This 30-minute interactive Q&A will cover several topics, including: * Why data pre-processing is important
* An analysis of RedPajama and SlimPajama
* What Cerebras did to pre-process this massive dataset
* How this impacts organizations like yours
LLMS
-
Data Preprocessing for Large Language Models: RedPajama Analysis
By
–
-

Cerebras Tech Talk: Dataset Processing for LLM Training
By
–
Join us for our next Cerebras Tech Talk! On Wednesday, June 28th at 11:00 AM PT, we will discuss how we cleaned and processed one of the largest public datasets so others can train higher-quality large language models. See next tweet for info. Register: https://
hubs.li/Q01VxRqb0 -
Language Models Risks Require Deeper Scientific Study
By
–
At the same time, we also find that LMs applied in this context pose risks that require (and illuminate areas for) deeper study.
-
Language Models for Digital Town Hall Synthesis
By
–
We find evidence that LMs have promising potential to help human facilitators and moderators synthesize the outcomes of online digital town halls—a role that requires significant expertise in quantitative & qualitative data analysis, the topic of debate, and writing skills.
-
Language Models Enhance Pol.is Platform for Diverse Dialogue
By
–
We collaborated with @compdem to research the opportunities and risks of augmenting the http://
Pol.is platform with language models (LMs) to facilitate open and constructive dialogue between people with diverse viewpoints. -
CEO discusses future of AI with custom LLMs and open source
By
–
Our CEO @andrewdfeldman joins @bigdata on
#TheDataExchangePod to discuss the future of #AI with custom #LLMs, Cerebras-GPT and how open source, efficient → the new era of AI Tune in here: -
LLaMA Innovation: Quantization and LoRA Techniques for Developers
By
–
But that sometimes breeds innovation—like in the case of LLaMA. And scrappy developers without access to massive training clusters like Google or OpenAI have to find ways around the problem. That's led to the widespread adoption of techniques like quantization and LoRA.
-
Inflection Announces Inflection-1 LLM Outperforming GPT-3.5
By
–
We’re proud to announce Inflection-1, the best-in-class LLM developed at Inflection! Inflection-1, which powers http://
Pi.ai, outperforms GPT-3.5, Chinchilla, and LLaMA on a number of academic benchmarks. More details in our technical memo: -
Top AI Experts from Major Tech Companies Join Our Team
By
–
Our small team includes some of the very best AI experts and engineers, who previously built the largest LLMs at DeepMind, Google, Microsoft, OpenAI and Meta. It’s down to this team that we outperform other companies despite being a year old. Join us!
-
Inflection-1 LLM Outperforms GPT-3.5 and Llama on Benchmarks
By
–
We have amazing results to announce! Inflection-1 is our new best-in-class LLM powering Pi, outperforming GPT-3.5, Llama and PALM-540B on major benchmarks commonly used for comparing LLMs.