Join us on Wednesday, May 8th, as NVIDIA's Richard Wang delves into #LLM training utilizing NVIDIA NeMo on @RedHat OpenShift, powered by @OracleCloud Infrastructure. Don't miss out. #RHSummit https://
redhat.com/en/summit
LLMS
-

NVIDIA NeMo LLM Training on OpenShift Infrastructure
By
–
-

Batch API Now Supports Larger Batches Multimodal
By
–
Bigger batches and more multimodal support in Batch API: https://t.co/kIybJPkwvO
— Greg Brockman (@gdb) 30 avril 2024Bigger batches and more multimodal support in Batch API:
-
LangSmith RAG Evaluation Video Series Launch
By
–
LangSmith Evaluations: RAG Evaluation (Document Retrieval)
— LangChain (@LangChain) 30 avril 2024
Evaluations can accelerate LLM app development, but it can be challenging to get started. We've kicked off a new video series focused on evaluations in LangSmith.
This is the 14th video in our series focusing… pic.twitter.com/P1qVbnJxm8LangSmith Evaluations: RAG Evaluation (Document Retrieval) Evaluations can accelerate LLM app development, but it can be challenging to get started. We've kicked off a new video series focused on evaluations in LangSmith. This is the 14th video in our series focusing
-
Efficient fine-tuning enables GPT-2 to approach GPT-4 performance
By
–
Most likely explanation for gpt2-chatbot: OpenAI has been working on a more efficient method for fine-tuning language models, and they managed to get GPT-2, a 1.5B parameter model, to perform pretty damn close to GPT-4, which is an order of magnitude larger and more costly to
-
Apple Develops Optimized AI Model Using Stanford Google Research
By
–
Apple builds a slimmed-down AI model using Stanford, Google innovations The phone giant's open-source large language model beats previous models by melding the insights of many researchers. https://
zdnet.com/article/apple-
builds-a-slimmed-down-ai-model-using-stanford-google-innovations/
… @Apple -
Using Claude, GPT-4, Haiku, Llama and Phi for AI workflows
By
–
Opus via LLM for CLI stuff, GPT-4 via ChatGPT for Code Interpreter and general convenience Haiku as an API for some features I'm building, Llama 3.8B and Phi-3 for running in my laptop I'm starting to try http://
meta.ai and Google Gemini for search-backed queries -
AGI Builders Meetup: Full-stack LLM Enterprise Use Cases
By
–
Please join us at the AGI Builders Meetup in SF on Tuesday, April 30th, https://
lu.ma/agi0430?tk=r9Z
mkb
…. It will be a lot of fun:
6:10 pm – 6:40 pm: Full-stack LLM for Enterprise Use Cases by Lucy Park, co-founder & CSO, Upstage -
Groq CEO discusses fast token generation and AI expansion plans
By
–
Chatting with @GroqInc
’s CEO @JonathanRoss321
. Groq has super fast token generation capabilities now. And, I was excited also to hear about his plans to scale up capacity aggressively and also expand this to other models than just LLMs! This is a good time to be building AI -
LMSYS Blind Test Mode Policy Documentation and Guidelines
By
–
I hadn't seen this before: LMSYS have clear documentation of their policies around this kind of "blind test mode" on their site: https://
lmsys.org/blog/2024-03-0
1-policy/#our-policy
…