24 hours for Groq and #Llama3. Read more about today's developments at https://
groq.link/llama3blog.
GENERATIVE AI
-
Groq and Llama3: 24 Hours of Development Updates
By
–
-
Lightning Studio: Try AI Development Platform Today
By
–
Try it on a Lightning Studio: http://
Lightning.ai -
400B Model Expected to Surpass GPT-4 Performance
By
–
they were “only” trying to match gpt4, and now 400b is likely to beat it
-
Medical Chatbot Deployment Risks and WHO Brand Responsibility
By
–
Yeah, in that case I'd question the wisdom of even attempting to release a chatbot! The scope of medical-related questions people could ask that is effectively infinite, I can't even imagine a QA process that would be thorough enough to responsibly release that with the WHO brand
-

llm-gpt4all 0.4 Release Notes Published
By
–
Release notes for llm-gpt4all 0.4 are here: https://
github.com/simonw/llm-gpt
4all/releases/tag/0.4
… -

Handling Chatbot Self-Help Queries with RAG and Prompting
By
–
If you ship a chatbot, it's now indisputable that people are going to ask it questions about how to use it Handling this isn't trivial but it's not unsolvable either – use RAG, fine-tuning or a meticulous system prompt, it's important to anticipate this use-case
-
GPT-2 Activation Memory and GPU Cache Analysis
By
–
Makes sense, in GPT-2 (124M) case we're currently doing B=4, T=1024, C=768 => 3M activations @ float32 => 12MB. A100 L2 cache is 40MB, and even L1, at 192KB/SM with 108 SMs => ~= 20MB (wow, that's more than I expected). The pleasures of smaller networks and caches…
-
Anthropic increases Haiku tokens amid API provider competition
By
–
5-10x times more Haiku tokens per day from @AnthropicAI – hard not to assume this is in reaction to the flood of new API providers selling cheap access to Llama 3
-
Flow Engineering Webinar Recording with LangChain Leaders
By
–
The recording from our Flow Engineering webinar with @hwchase17 and @itamar_mar is up! https://
youtube.com/watch?v=eBjxz7
qrNBs
… "Flow Engineering" is a term that has been gaining in popularity recently. The first time it was mentioned as term was in CodiumAI paper on AlphaCodium, where they -
Stable Diffusion 3 and SD3 Turbo Now Available on Poe
By
–
Stable Diffusion 3 and the faster SD3 Turbo are hosted by @FireworksAI_HQ and available at SD3 https://
poe.com/StableDiffusio
n3
…, and SD3-Turbo https://
poe.com/SD3-Turbo, and across all Poe apps. We look forward to seeing what you create! (2/2)