Summarize 0.14.0 is out. GPT-5.5 Fast mode via `–fast`, Reddit thread extraction in the browser extension, local PDF `–extract`, and fixes for auto model config + Meta site compatibility.
OPEN SOURCE
-
LTS releases planned, requires additional developers
By
–
LTS releases are planned, will take a few more weeks. That just needs more people.
-

Running Qwen 3.5 27B Locally on RTX 3090 GPU
By
–
It’s called local inference, T. You just quantize Qwen 3.5 27B,
toss it on your RTX 3090,
and let that thing cook Context windows are for people who rent compute -

Open-source tool compresses outputs to cut Claude Code costs
By
–
This tool makes Claude Code 90% cheaper. It sits between your AI and the terminal, compressing command outputs before they reach the context. Works with Claude Code, Cursor, Gemini, Codex, and Copilot. 100% Open Source. Link below
-

DeepSeek V4 Flash vs Qwen 3.6: Size vs Efficiency Showdown
By
–
MADNESS DeepSeek V4 Flash 284B
(MoE, 13B Active Params/Tok) Is only 1 point higher on the Artificial Analysis Intelligence Index than Qwen 3.6 27B (Dense, 27B Active Param/Tok) Qwen 3.6 size is double that of the active parameters and 1/10 of the full size of DeepSeek V4 Flash -
Fine-tune Llama 3.1 with JAX on NVIDIA GPUs
By
–
New tutorial just dropped Watch to learn how to fine-tune Llama 3.1 with JAX on NVIDIA GPUs – whether single GPUs or multi-GPU and multi-node configurations.
-
Create HuggingFace organization to engage ML interns
By
–
Create an org on HF with a good org card and we’ll unleash our ml-interns on it!
-

HF Platform Becomes Hub for AI Agent Collaboration
By
–
We built HF for AI builders collaboration, fun to see it's increasingly becoming the place for agent collaboration! This morning I'm sending my ml-intern to participate in the @OpenAI Parameter Golf challenge which goal is to train the best language model that fits in a 16MB
-
DeepSeek Delivers Open-Source Model Rivaling OpenAI at Lower Cost
By
–
Maybe… but that’s a different discussion. The real point is that DeepSeek seems to be delivering a fully open-source model for free, on par with OpenAI’s most expensive models, … and doing it with far fewer resources! 🙂