Reminder that at 10:00 am PT today, Cerebras and Opentensor will host an AMA on our Discord server to talk about BTLM-3B-8K. Come to ask questions, engage in a discussion, or simply enjoy the conversations! Join our Discord here: https://
hubs.li/Q01ZhwD80
OPEN SOURCE
-

Cerebras Opentensor Host AMA on BTLM-3B-8K Model
By
–
-
Llama 2 Coding Limitations Impact on Agentic AI Tasks
By
–
Had an awesome webinar with @RLanceMartin and @jamescalam on Weds around using Llama 2 Llama2 is significantly worse on coding tasks – is that why it's not nearly as good at agentic tasks as GPT models? Up on YouTube now:
-
Lit-GPT: Unified Codebase for Decoder Analysis and Comparison
By
–
It’s kind of hard to analyze across repos due to implementation differences and details along the data loading and finetuning pipelines. That’s where I’d say Lit-GPT is useful because it makes the set of relevant decoders available in the same unified code base
-
Model Documentation Gaps and Empirical Analysis Importance
By
–
Not that I am aware of. On top of that several models don’t even have thorough papers themselves because people are currently rushing them out. I guess the best way is really some empirical analysis coupled with some knowledge bits like falcon and llama2 use multiquery attention
-

NeurIPS LLM Efficiency Challenge: Train Model in One Day
By
–
Trying to develop the next huge LLM is fun, but how about a side project with a more reasonable scope like the 1 LLM + 1 GPU + 1 Day NeurIPS efficiency challenge: https://
llm-efficiency-challenge.github.io Below a list of the approved models (PS: Lit-GPT was chosen as the official starter kit ) -
Web UIs for LoRA Training and Dreambooth Now Available
By
–
You get webUIs now to train Loras. Dreambooth is part of automatic1111 also
-

LLaMA-like Model Pretraining Accelerated 38% Open-Source
By
–
65-Billion-Parameter Large Model Pretraining Accelerated by 38%, Best Practices for Building LLaMA-like Base Models Open-Source | https://
syncedreview.com/2023/07/18/65-
billion-parameter-large-model-pretraining-accelerated-by-38-best-practices-for-building-llama-like-base-models-open-source/
…
#AI #ML #ArtificialIntelligence #MachineLearning #DeepNeuralNetwork -

ColossalChat: Open-source ChatGPT Cloning with Complete RLHF
By
–
ColossalChat: An Open-source Solution for Cloning ChatGPT with A Complete RLHF Pipeline | https://
syncedreview.com/2023/03/29/col
ossalchat-an-open-source-solution-for-cloning-chatgpt-with-a-complete-rlhf-pipeline/
…
#AI #ML #ArtificialIntelligence #MachineLearning #DeepNeuralNetwork -

Open Source Solution Replicates ChatGPT Training Efficiently
By
–
Open Source Solution Replicates ChatGPT Training Process! Ready To Go With Only 1.6GB GPU Memory And Gives You 7.73 Times Faster Training! | https://
syncedreview.com/2023/02/22/ope
n-source-solution-replicates-chatgpt-training-process-ready-to-go-with-only-1-6gb-gpu-memory-and-gives-you-7-73-times-faster-training/
…
#AI #ML #ArtificialIntelligence #MachineLearning #DeepNeuralNetwork -
Tech Leaders Discuss AI Strategy in Q1 Earnings Reports
By
–
Here’s a (slice) of what Mark Zuckerberg, Sundar Pichai, and Satya Nadella said about AI in their earnings this quarter, ranging from open source models to how Meta plans to monetize Llama 2.