What you can expect from the Databricks RAG Application: A vector search service to power semantic search Online feature and function serving for structured data Fully managed foundation models Learn more: https://
bit.ly/3uUuuLR
LLMS
-

Databricks RAG Application: Vector Search and Foundation Models
By
–
-

Master ChatGPT Prompts with Practical Cheat Sheet Guide
By
–
Master the art of #ChatGPT prompts with this clever cheat sheet! Whether you're drafting essays, creating code, or designing a Gantt chart, it's a handy guide for any role.
Level up your #AI game by following @ingliguori for more insights on leveraging tech for success. -
Groq’s Chat Platform Achieves 300 Tokens Per Second with Llama2
By
–
…and that open-source list means you could learn about the samely named model but on different platforms, yet it will perform differently. This is challenging for providers also. http://
Chat.Groq.com is the fastest for Llama2, 70B at ~300 tokens per second per user but that -
Groq Hosts AI Compute Cocktail Party at CES 2024
By
–
Some Grogsters, & CEO @JonathanRoss321
, will be @CES next week hosting a cocktail party. Will you be there? Want to come hang out and chat about the future of AI compute and how to ready your org? DM us and maybe we'll send you the invite form. 😉 #betterongroq #LLMs #AI #CES24 -
Groq LPU Engine: Kernel-free Llama 2 70B Inference Performance
By
–
Kernel-free, no Cuda, Compiler-only solution.
The current implementation, running on our LPU Inference Engine uses our own processor. The model is Llama 2, 70B 4k sequence at FP16…We haven't even hit the next gear through those other methods. Plenty in the tank left! -
OpenAI Custom RAG with LangChain: Chat with Videos
By
–
Building an OpenAI Custom RAG with LangChain: The Ultimate Tutorial to Chat with your Videos! Awesome tutorial by @austinbv on how to use Whisper + OpenAI + LangChain to talk to your videos Shows off transcription, LCEL, streaming
-

ColBERT LangChain Integration Guide Easiest Method
By
–
Want to use ColBERT with LangChain? This is the easiest way we know of
-
H2O.ai and NVIDIA Enable LLM Deployment for Financial Services
By
–
.
@h2oai and NVIDIA are working together to provide an end-to-end workflow for financial institutions, using NVIDIA AI Enterprise. Companies can develop and deploy their own #LLMs to power #generativeAI use cases in #financialservices. Read the blog here. -
Option pour envoyer requêtes dans menu GPT personnalisé
By
–
It is also expected that an option to send these requests will be soon available in the dropdown menu of a custom GPT itself. h/t @imrat https://
buff.ly/3REwwHS -
H2O.ai and NVIDIA Partner for Enterprise LLM Deployment
By
–
@h2oai and NVIDIA are working together to provide an end-to-end workflow for financial institutions, using NVIDIA AI Enterprise. Companies can develop and deploy their own #LLMs to power #generativeAI use cases in #financialservices. Read the blog here.