If you liked my NN+Gzip article last week (
https://
magazine.sebastianraschka.com/p/large-langua
ge-models-and-nearest
…), you'll probably love @abhi9u & @alepiad 's follow-up. This one looks at the NN+Gzip method from a compression algo perspective comparing Gzip, Huffman coding, LZ77/LZ4, and Bzip2: https://
codeconfessions.substack.com/p/lz77-is-all-
you-need/
…
LLMS
-
NN+Gzip Follow-up: Compression Algorithms Comparison Study
By
–
-
Llama 2-Chat Updates: Reducing False Refusals and Improving Sanitization
By
–
Hearing community feedback & following internal research & analysis, we've pushed two new updates to the Llama repo to reduce false refusal rates seen with Llama 2-Chat models & improve token sanitization. Full details https://
bit.ly/3QuRNER -

Groq Language Processing Units Transform the Future of Computing
By
–
Join us to see how @GroqInc Language Processing Units™ (LPUs) are changing the future of compute. http://
groq.link/gsaugust -

Llama 2 Model Comparison: 7B vs 13B vs 70B Guide
By
–
What's the difference between Llama 2 7b, 13b and 70b, and when should you choose one over the other? @zeke breaks it down for us: https://
replicate.com/blog/all-the-l
lamas
… -
Vision-Language Pre-training: Fundamentals, Advances, and Future Directions
By
–
Vision-Language Pre-training: Basics, Recent Advances, and Future Trends
-

Vision-Language Pre-training: Comprehensive Overview of Methods and Architectures
By
–
Vision-Language Pre-training(VLP): Basics, Recent Advances, and Future Trends This is one of the best and comprehensive resources that provides an overview of approaches(basic techniques, model architectures, tasks, benchmarks) used in vision-language pre-training. Just as
-

Deploy and Fine-tune Open-Source LLMs with Ludwig AI
By
–
Want to learn how to deploy and #finetune open-source #LLMs on your own data the easy way? Join our hands-on webinar to learn about the latest release of open-source @ludwig_ai and how to build custom LLMs in just a few commands. Save your spot: https://
pbase.ai/3s1KiLy -
Model Performance Comparison with 11B T5 Parameters
By
–
Probably about the same as the 11B T5 model
-
Open Source AI Checkpoint Released on Hugging Face
By
–
When we do a LikeRT human evaluation of our checkpoint and compare with other models, we achieve comparable results. We are open sourcing the checkpoint on @huggingface for the community to try. (8/10)
-
Open License Model Released for Long Sequence Understanding
By
–
It is available with an open license and created to understand long sequence capabilities. It is not meant to be a drop-in replacement for chat models. We are excited to see how people use the checkpoint. (9/10)