Congrats on the release! Love the open weights!
OPEN SOURCE
-

AI21 Launches Jamba 1.6: Top Open Model for Enterprise
By
–
Today we launched Jamba 1.6, the best open model for private enterprise deployment. AI21’s Jamba outperforms Cohere, Mistral and Llama on key benchmarks, including Arena Hard, and rivals leading closed models while maintaining unmatched speed and quality. Now available on
-
Alibaba QwQ-32B: Reasoning AI Matches DeepSeek-R1 Cheaper
By
–
Alibaba’s Qwen team dropped QwQ-32B, a reasoning AI that matches or surpasses DeepSeek-R1 at much less cost
— Rowan Cheung (@rowancheung) 6 mars 2025
—20x smaller than DeepSeek-R1
—Priced $0.20 per million input and output tokens
—Open-sourced under Apache 2.0
Here it is running on M4 Max:pic.twitter.com/J40JG0jGtCAlibaba’s Qwen team dropped QwQ-32B, a reasoning AI that matches or surpasses DeepSeek-R1 at much less cost —20x smaller than DeepSeek-R1
—Priced $0.20 per million input and output tokens
—Open-sourced under Apache 2.0 Here it is running on M4 Max: -
Alibaba Qwen 32B Full Release Matches DeepSeek R1 Performance
By
–
We’ve been hyped to support @Alibaba_Qwen 32B during its preview phase & now the full version is here, performing on par with #DeepSeek R1! Devs, get ready—our update drops soon! More info on @huggingface
-
HeadInfer: Long-Context LLM Inference on Consumer GPUs
By
–
HeadInfer: Unlocking Long-Context LLM Inference on Consumer GPUs (Million-level Tokens)
*long-context inputs require large GPU memory.
*A standard LLM like Llama-3–8B requires 207GB of GPU memory for 1 million tokens — far beyond the capabilities of consumer GPUs like the RTX -
LangGraph Java Development Announced
By
–
maybe buried the lede a bit – we're working on LangGraph Java 🙂 If that is interesting – get in touch
-

Claude Code Ports LangGraph to Java: Real World Application
By
–
good thread on using claude code to do a real world task (porting langgraph to java)
-

QwQ-32B Matches DeepSeek R1 at 20X Smaller Scale
By
–
DeepSeek just got DeepSeek'd. QwQ-32B is on par with DeepSeek R1 671B but ~20X smaller. AGI will run on your computer. Then it'll run on your phone.
-

QwQ-32B Powers Qwen Chat with Reasoning
By
–

Qwen has released a new reasoning model, QwQ-32B, which now powers Qwen Chat when you select Qwen2.5-Plus with Thinking (QwQ).
-
Hugging Face Libraries Surge: 25% Created in Early 2025
By
–
wow! >25% of Hugging Face libraries with the highest GitHub star counts (10k+) have been created in the first two months of 2025 alone staying relevant in the long run is the hardest skills in AI. each wave changes everything, our community keeps growing