those too, but i think more importantly i'm asking about the relationship between model parameters, training steps (whatever that means in the embeddings case), and downstream performance
LLMS
-
Scaling Laws for Language Models Explained
By
–
no, scaling laws like the scaling laws for language models:
-
Scaling Laws for Embedding Models: Research Initiatives
By
–
who's working on scaling laws for embedding models? could be any of:
• text embeddings (DPR, GTR, GTE…)
• image embeddings (SimCLR, DINO…)
• recommendation systems (?)
• multimodal embeddings (CLIP, ImageBind…)
• any other type of embeddings… -
GPT-3.5 includes text-davinci-002 and code-davinci-002 models
By
–
Officially, GPT-3.5 includes text-davinci-002 and code-davinci-002 even though neither were branded as 3.5 until after text-davinci-003 was released. Full list: https://
platform.openai.com/docs/models/gp
t-3-5
… -

AI21 Labs at Vegas Event: Live Demos and LLM Innovation
By
–
We’re here in Vegas! Swing by booth 205 to catch the action—live demos, cool swag, 1:1 sessions, and see how we're shaping the future of LLMs.
-

Calibrated LLMs must hallucinate; RLHF reduces calibration
By
–
“Calibrated Language Models Must Hallucinate” — Hallucination rate from a pre-trained LLM ≈ proportion of facts seen only once in training. RLHF reduces hallucination but makes LLMs less well calibrated as models. Real text has news; LLMs need to unlearn that fact to please us.
-

DLT Deepnote Analysis Women’s Wellness Violence Trends
By
–
DLT & Deepnote in women's wellness and violence trends: A Visual Analysis https://
bit.ly/3sIg0yh
#AI #MachineLearning #DeepLearning #LLMs #DataScience -

Scale AI and NVIDIA Collaborate on HelpSteer LLM Dataset
By
–
.
@scale_AI worked with @nvidia on HelpSteer, a first-of-kind multi-attribute preference dataset across correctness, coherence, complexity & verbosity. We look forward to more research collaboration with @nvidia on LLMs! Check out the paper here: https://
arxiv.org/abs/2311.09528 -
Reducing LLM Hallucinations Through Confidence Detection
By
–
Why Hallucinations happen with LLMs?
— Louis-François Bouchard 🎥🤖 (@Whats_AI) 27 novembre 2023
Paige Bailey discusses how language models are less likely to have hallucination. She also covers a nice approach where models are able to spot when it has low confidence in answering the question or ask follow up question to ensure… pic.twitter.com/eLzhVZ4TBYWhy Hallucinations happen with LLMs? Paige Bailey discusses how language models are less likely to have hallucination. She also covers a nice approach where models are able to spot when it has low confidence in answering the question or ask follow up question to ensure