If the initial benchmarks scores (and graphs used for PR) showcased a 3x reduction in size for the same performance, I think the broader public reception would have been less tepid. Just looking at this, it just seems 7% behind other existing models in the 4.x series…
LLMS
-

Hugging Face Releases Transformers Version 5 RC
By
–
Today we release the transformers version 5 RC! With this, we enable e2e interoperability with our friends in ecosystem, ease up adding new models and simplify the library Read our blog to learn more: https://
huggingface.co/blog/transform
ers-v5
… -
Kling AI Unveils Multimodal Kling O1 Model
By
–
Kling AI released the Kling O1 model with multimodal understanding that can unify input across texts, images, and videos!
— 🚨 AI News | TestingCatalog (@testingcatalog) 1 décembre 2025
AI video week 🔥 https://t.co/r7WXgVtLRf pic.twitter.com/HtXMzX5ZuhKling AI released the Kling O1 model with multimodal understanding that can unify input across texts, images, and videos! AI video week
-
Opus 4.5 Model Review: Improved Writing and Coding Capabilities
By
–
Opus 4.5: 7.5-8/10 helpful. I finally trust this model to write for me and it actually has good judgement/taste as to what matters. For coding, it feels like it can just work forever and not get stuck in the same vibe coding doom loops as previous models. Some things are
-

DeepSeek-V3.2 outperforms GPT-5 on benchmarks after release
By
–
DeepSeek-V3.2 dropped last night and absolutely crushes GPT-5 on benchmarks
-
Post-training bottlenecks solved through refined methods and data
By
–
"The lesson is post-training bottlenecks are solved by refining methods and data" Zhibin Gou (@zebgou) If Gemini-3 proved continual scaling pretraining, DeepSeek-V3.2-Speciale proves scaling RL with large context. We spent a year pushing DeepSeek-V3 to its limits. The lesson is post-training bottlenecks are solved by refining methods and data, not just waiting for a better base. — https://nitter.net/zebgou/status/1995462720078934213#m
-

DeepSeek Releases DeepSeek-V3.2 and V3.2-Speciale Models
By
–
DeepSeek released DeepSeek-V3.2 and DeepSeek-V3.2-Speciale models with a top tier performance according to the benchmarks! Upgrade time
-
Claude Opus 4.5 Launches with Advanced Agentic Features
By
–
Claude Opus 4.5
Solves complex bugs and multi-doc analysis; excels in legal/enterprise reasoning; launched on Microsoft Foundry with agentic features. https://
buff.ly/Hqv6I3c -

Baidu ERNIE 5.0 Surpasses GPT-5 on Key Benchmarks
By
–
6️⃣ Baidu introduced ERNIE 5.0, a multi-modal model that reportedly surpasses GPT-5 on key benchmarks across text, image, audio, and video. pic.twitter.com/xwR2GtXkBx
— Futurepedia – Learn to Leverage AI (@futurepedia_io) 1 décembre 2025Baidu introduced ERNIE 5.0, a multi-modal model that reportedly surpasses GPT-5 on key benchmarks across text, image, audio, and video.
