but also, I'm pretty confident that anytime the leading open-source model seat will stay vacant for 2+ months a new (or old) team will rise to take it. Open-sourcing a new SOTA model and sometimes –like DeepSeek did– they might over-shoot their goal and accidentally take the
OPEN SOURCE
-

DeepSeek-R1 Open Source Model Available as NVIDIA NIM Microservice
By
–
The open source DeepSeek-R1 model is now available as an NVIDIA NIM microservice preview on http://
build.nvidia.com to help developers securely experiment with its advanced AI reasoning capabilities. -
Dario’s DeepSeek Essay: Closed-Source Justification Critique
By
–
Finally took time to go over Dario's essay on DeepSeek and export control and to be honest it was quite painful to read. And I say this as a great admirer of Anthropic and big user of Claude* The first half of the essay reads like a lengthy attempt to justify that closed-source
-

SambaNova Partners with HuggingFace for 10x Faster AI Inference
By
–
ICYMI: We've partnered with @HuggingFace to bring 10x faster #AI inference speeds to devs! @AI at Meta's Llama 3 & @Alibaba_Qwen models SambaNova Cloud's Llama Guard & Qwen QwQ Easy integration with minimal code changes @deepseek_ai coming soon! Try it now
-
DeepSeek Shifts: China Advances, Open Models Commoditize AI
By
–
The buzz over DeepSeek this week crystallized, for many people, a few important trends that have been happening in plain sight: (i) China is catching up to the U.S. in generative AI, with implications for the AI supply chain. (ii) Open weight models are commoditizing the
-

DeepSeek R1 70B Outperforms GPT-4o and o1-mini
By
–
DeepSeek’s R1 70B combines the powerful reasoning ability of the full R1 model with the size and speed of Llama 70B. R1 70B outperforms GPT-4o and o1-mini across a range of general and reasoning benchmarks, making it the most capable Llama 70B variant by far.
-
Button switches model to o1
By
–
This button changes the model to o1 so the response you see is generated by o1 too
-

Mistral AI releases Mistral Small 3
By
–
Mistral AI announced Mistral Small 3, a latency-optimized 24B-parameter model released under the Apache 2.0 license.
-

Open-source AI tool challenges Google NotebookLM
By
–
This open-source #AI tool was built in a day and it’s coming for Google’s NotebookLM
by @MichaelFNunez @VentureBeat Read more: https://
buff.ly/3XPfppF #ArtificialIntelligence #MachineLearning #ML #Technology cc: @terenceleungsf @alvinfoo @bernardmarr -
New 24B Open-Source AI Model Achieves 81% MMLU Performance
By
–
A new model to hasten AI progress. 24B, 81% MMLU, no RL for now! We're super excited to see the latest development in international open-source AI (kudos to Deepseek!), and cannot wait to bring new contributions to it. We're renewing our commitment to using Apache licenses. AI