Claude 3.5 Sonnet knows its own version remarkably well — not only is it right, it’s aware of the unique ambiguity of “3.5” (many users unofficially call this release 3.5.1 or 3.6) and gives the release month. (Claude 3.5 Haiku does misidentify as Claude 3 Haiku, though)
LLMS
-

Comparing LLMs on identifying their version
By
–

This is such a vibes-based eval, but the first prompt I give any new LLM is “Which version is this?” and DeepSeek-V3 nailed it See below for how Claude, Gemini, ChatGPT, and Grok fare on the same — TLDR: it’s all over the map
-

Llamathon Hackathon: Win Credits and Dev Tiers
By
–
Get ready to run the Llamathon! 🦙
— SambaNova (@SambaNovaAI) 26 décembre 2024
Join our Llama-thon #Hackathon with @Gradio and get a chance to win:
🤗 $10k in @HuggingFace credits
🚀 Free #dev tiers on SambaNova
… and more!
See details below 👇#AIGet ready to run the Llamathon! Join our Llama-thon #Hackathon with @Gradio and get a chance to win: $10k in @HuggingFace credits Free #dev tiers on SambaNova … and more! See details below #AI
-

Data LLMs: Extracting Insights and Stories from Your Data
By
–
Your data has a story to tell. You can use @AbacusAI #DataLLM to detect the patterns in your data, extract hidden insights, and generate data stories, reports, #DataViz, & narratives for you: https://
blog.abacus.ai/blog/2023/08/2
4/data-llm-get-insights-from-your-data/
…
————
#AI #LLMs #MachineLearning #DataScience #GenerativeAI #IoT -
DeepSeek’s 5.5M USD Breakthrough: Frontier AI Cost Revolution
By
–
Agreed. On a similar note, what are your thoughts on the recent DeepSeek release? The frontier *just* costs 5.5M USD (+ a cracked team) – and I'm here for it!
-
FP8 Precision Format Discussion in Machine Learning
By
–
My understanding from the paper is that it's fp8. Maybe @zizhpan can confirm
-
Text Generation Inference v3.0.1 release fixes reported issue
By
–
this should not happen, can you try with the latest tagged image – 3.0.1 happy to flag it to the team if it still doesn't work! sorry for the inconvenience! https://
github.com/huggingface/te
xt-generation-inference/releases/tag/v3.0.1
… -

Frontier Model Development Cost Drops to 5.5 Million USD
By
–
Scarcity breeds Innovation – cost to build a frontier model – 5.5 Million USD In a way, it's the maximum it'd be (Note: H800s have ~2x slower chip-to-chip data transfer) This cost, will only go down further and further as we continue to find newer walls to scale!
-
Qwen and Meta GPU Mobilization in AI Competition
By
–
I wouldn't discount Qwen or Meta either – at least the latter is mobilising metric fk ton of GPUs
