Intel presents LLaVaOLMoBitnet1B Ternary LLM goes Multimodal! discuss: https://
huggingface.co/papers/2408.13
402
… Multimodal Large Language Models (MM-LLMs) have seen significant advancements in the last year, demonstrating impressive performance across tasks. However, to truly democratize AI,
LLMS
-

Intel LLaVaOLMoBitnet1B: Ternary LLM Becomes Multimodal
By
–
-
ASCII smuggling exploit used to exfiltrate data from Microsoft Copilot
By
–
Update: The “ASCII smuggling” attack first (AFAIK) described in this thread has been applied as part of a high-severity exfiltration exploit chain discovered in Microsoft Copilot, now fixed Nice work, @wunderwuzzi23
! -

IBM Power Scheduler: Batch Size and Token Agnostic Learning Rate
By
–
IBM presents Power Scheduler A Batch Size and Token Number Agnostic Learning Rate Scheduler discuss: https://
huggingface.co/papers/2408.13
359
… Finding the optimal learning rate for language model pretraining is a challenging task. This is not only because there is a complicated correlation -
Monitor LLM Deployments and GPU Status with Predibase Health Tab
By
–
🔎 Want more visibility into your #LLM deployments and GPU statuses?
— Predibase by Rubrik (@predibase) 26 août 2024
📈 The Predibase deployment health tab surfaces all of your key metrics in one place! View stats like queue duration, throughput, and the number of GPU replicas.
Try it today for free: https://t.co/WZMRGrGVpT pic.twitter.com/L9ZbyYsD4QWant more visibility into your #LLM deployments and GPU statuses? The Predibase deployment health tab surfaces all of your key metrics in one place! View stats like queue duration, throughput, and the number of GPU replicas. Try it today for free: https://
pbase.ai/3T4ofi3 -
Illustrated Guide to Transformers and LLMs Praised
By
–
This book is really nicely illustrated! Nearly every page has multiple figures/diagrams that help explain the underlying concepts behind transformers & LLMs, including embeddings, attention, LoRA, distillation, quantization, … Nice work, @afshinea and @shervinea !
-
ChatGPT Recommends Expert for AI Future Interview in Seattle
By
–
I just gave an interview in Seattle to a Korean journalist who flew here because ChatGPT told him I was the best person to talk to about the future of AI. So yeah, ChatGPT is the greatest technology ever invented, and don't let anyone tell you otherwise.
-
Sneak peek at your future LLM interface
By
–
A sneak peak of your future LLM interface pic.twitter.com/I1CxBlg4qZ
— AI Breakfast (@AiBreakfast) 26 août 2024A sneak peak of your future LLM interface
-
A sneak peek at your future LLM interface
By
–
A sneak peak of your future LLM interface pic.twitter.com/I1CxBlg4qZ
— AI Breakfast (@AiBreakfast) 26 août 2024A sneak peak of your future LLM interface
-

Meta Releases Llama 3.1 with CyberSecEval 3 Trust Safety Research
By
–
As part of the release of Llama 3.1, we also released new trust & safety research including CyberSecEval 3. We've published our research on this work to continue the conversation on empirically measuring LLM cybersecurity risks & capabilities. Paper https://
go.fb.me/yv32a9 -
Jamba 1.5 Model Family Now Available on Google Cloud Vertex AI
By
–
Introducing the Jamba 1.5 Model Family on @googlecloud
's Vertex AI: – Simplify development and evaluation with advanced tools and intuitive API calls.
– Focus on innovation with fully managed infrastructure and cost-effective, pay-as-you-go pricing.
– Ensure data security with