Thanks @togethercompute – we're exciting to be making our Jamba models available on Together AI soon!
LLMS
-
Exploring Native Multimodal and Long Context AI Capabilities
By
–
It's supposed to be native now Include other usecases "Test out long context, native multi-modal (image, video and audio), structured outputs, and more!"
-
LLMs Limited to Simple Problems Without Deep Expertise
By
–
True, only for relatively simpler problems, that do not require deep domain expertise and layered, concentrated thinking. The time spent on frivolous 15 minute shallow groupthink is disproportionately high.
-
New Free Course on Improving LLM Applications Accuracy
By
–
New course on @DeepLearningAI
: Improving Accuracy of LLM Applications https://
go.fb.me/zfwvd8 Created in collaboration with DL, Meta & @LaminiAI
, this free course covers topics like evaluation frameworks, instruction & memory fine-tuning, LoRA + training data generation. -

Knowledge Distillation: Creating Efficient Smaller AI Models
By
–
Distillation: making smaller, stronger models since 2015… @OriolVinyalsML @geoffreyhinton https://
arxiv.org/abs/1503.02531 -
AI Models Progressing Rapidly: Expert Shows the Ropes
By
–
Looks awesome! Happy to show you the ropes! Crazy how these models are progressing.
-
LLMs start rickrolling people as step towards AGI
By
–
Today's step towards AGI: LLMs start rickrolling people.
-
OpenAI Offers Free GPT-4o Fine-Tuning
By
–
[#Article]
@OpenAI launches GPT-4o fine-tuning with limited free offer https://actuia.com/actualite/open-ai-lance-le-fine-tuning-de-gpt-4o-avec-une-offre-gratuite-limitee/
… #AI #ArtificialIntelligence -
Opinion: RLHF vs the ‘Base Model’ Myth
By
–
People hear RLHF and think censorship, so they expect the base model to be untainted, honest, free — all those nice Grokian promises. And it is! But it’s also helplessly insane, and lost in the delusion the entire world is improv theater. “Base model” — not “based model.”
-

Fine-tuning drives record AI performance on SWE-bench and BIRD-SQL
By
–

Success stories are already rolling in. Cosine used fine-tuning to set a new record on the SWE-bench benchmark with its AI agent Genie. Distyl topped the BIRD-SQL benchmark with impressive accuracy in SQL tasks.