The amount of non-deterministic behavior LLMs introduce makes me believe in techno-sorcery futures. Let's chant a verse, burn candles, perform data prayers and result purification rituals. Anything to get this endpoints perform predictable across different input
LLMS
-
GPT-2 Reproduction with Increased Channel Size and Memory Optimization
By
–
We want to do a full GPT-2 repro, at channel size 1600 this is 2.1X higher C. And we'll want to ~max out batch dim to fit in memory too. So the "easy times" will be over soon.
-

Understanding AI Layers: From Broad Concepts to Generative AI
By
–
Peeling the layers of #AI: From broad AI concepts to the niche realm of #GenerativeAI, each layer builds upon the next. Stay updated with @ingliguori for insights into AI's complex structure and master its potential with 'The Digital Edge' https://
bit.ly/3u4pILl #DeepLearning -

Novel One-Sentence Startup Pitch Structure Across AI Models
By
–
Here's an interesting question to use to compare models: "GPT-4, Llama 3, Claude 3, Gemini 1.5, give me a novel structure for a one-sentence startup pitch and teach me how to use it"
-

Document Chunking Techniques for Better RAG Applications
By
–
But, How is Chunking Done? To create Better RAG applications, you need to know how to split or chunk the documents so you preserve the content while asking questions. Data Science Basics shows how to do this with LangChain https://
youtube.com/watch?v=tMwdl9
hFPns
… -
Llama 3: Dense Model Efficiency vs Sparse MoE Scaling Strategy
By
–
Thing that most impresses me about Llama 3: how did they pack so much knowledge and reasoning into a dense 8b and a 70b so well, when everyone else has been scaling sparse MoEs. This still doesn’t mean having a lot of GPUs is not important. Probably even more important
-
Groq’s AI Chip Achieves 800 Tokens Per Second Performance
By
–
https://
venturebeat.com/ai/groqs-break
through-ai-chip-achieves-blistering-800-tokens-per-second-on-metas-llama-3/
… Thanks @VentureBeat for bringing the story to light. We really want to advance the landscape for AI Inference and the infrastructure necessary to usher in real-time AI for developers/applications. More to come… -
8B Model Enables Creation of Diverse AI Experiences
By
–
8b is so good. Can create a lot more experiences with it. We have some ideas. Stay tuned!

