CodeLlama — a version of Llama2 that was fine-tuned for code tasks is live now. Available in 7B, 13B and 34B. https://
ai.meta.com/blog/code-llam
a-large-language-model-coding/
…
OPEN SOURCE
-

CodeLlama: Meta’s Llama2 Fine-Tuned for Code Tasks
By
–
-

Hugging Face Reaches $4.5B Valuation in New Funding Round
By
–
Hugging Face, the open source darling that has become pretty much the home of all OSS LLM and Diffusion models, has hit a $4.5 billion valuation with a new funding round. I spoke with Clement Delangue about what it means and how they balance enterprise plans and community growth.
-
Stability AI Achievement Recognition and Collaborative Diffusion Initiative
By
–
Exceptional Work, Congratulations!#diffusetogether #stabilityai https://t.co/mwTTzTh4Rs
— Stability AI (@StabilityAI) 24 août 2023Exceptional Work, Congratulations!
#diffusetogether #stabilityai -
llama.cpp inference and gptq quantization techniques exploration
By
–
Oh, reading a bit more about llama.cpp (
https://
github.com/ggerganov/llam
a.cpp
…), that's only inference, not training? I haven't tried since I don't have the model checkpoints on my laptop, but you may be able to use gptq.int4 quantization then: https://
github.com/Lightning-AI/l
it-gpt/blob/main/tutorials/quantize.md
… -
QLoRA 4-bit NormalFloat format supported only on Nvidia GPUs
By
–
Ah sorry, I meant M1/M2 chips (not specifically M1/2 CPUs). As far as I know, the 4-bit NormalFloat format that is used in QLoRA is currently only supported on Nvidia GPUs (
https://
github.com/TimDettmers/bi
tsandbytes/issues/485
…). Maybe the repo you mentioned uses a different type of quantized training. -

LoRA Compared to Llama-Adapter and Llama-Adapter v2
By
–
LoRA is a parameter-efficient finetuning technique, yes. I recently compared to Llama-Adapter and Llama-Adapter v2:
-

SeamlessM4T Breakthrough in Multilingual Speech Translation
By
–
SeamlessM4T represents a significant breakthrough in the field of speech-to-speech & speech-to-text by addressing the challenges of limited language coverage & a reliance on separate systems. More details https://
bit.ly/45g2pMq -

Fine-tuning LLaMA2 Workshop: One-Day In-Person Event
By
–
If you want to explore finetuning LLaMA2, we'll be talking at and helping out with a 1 day in-person event focused explicitly on finetuning OSS models Hopefully this guide will come in handy! RSVP here (s/o @swyx and @NaderLikeLadder for organizing): https://
partiful.com/e/T4ngRPaU2uUT
XM8pN17d
… -
Flowise v1.3.4 Release: VectorDB and SQLChain Enhancements
By
–
It's release time again Flowise v1.3.4 is now freshly baked with: VectorDB Document Loader Chat history variable 3 new vector stores: Milvus, PgVector, Zep SQLChain now supports MySQL, MsSQL, PostgreSQL Serp API Loader
-
Lit-GPT: Lightning AI’s New GitHub Repository Launch
By
–
Lit-GPT 🙂 https://
github.com/Lightning-AI/l
it-gpt
…