Oh yeah you probably will. I am currently running tons of experiments to decide which of the Olmo 3 and DeepSeek V3.2 GRPO tweak I should add in the final version.
No side objective really except I'd do an honest evaluation to make sure that the model actually learns well.
LLMS
-
Experimenting with Olmo 3 and DeepSeek V3.2 GRPO tweaks
By
–
-
Gemini 3 Flash Enables Research Paper Analysis and Comparison
By
–
Introducing Gemini 3 Flash for understanding research papers 🚀
— alphaXiv (@askalphaxiv) 17 décembre 2025
Highlight any section of a paper to ask questions and “@” other papers for quick context, comparisons, and benchmark references pic.twitter.com/u3pZc78mn5Introducing Gemini 3 Flash for understanding research papers Highlight any section of a paper to ask questions and “@” other papers for quick context, comparisons, and benchmark references
-

Model Convergence Shifts AI Differentiation to Data Foundations
By
–
On @TBPN
, Databricks CEO and co-founder @alighodsi joined @johncoogan and @jordihays to discuss how model capabilities are converging, and why differentiation increasingly comes from strong data foundations and reasoning over proprietary data. They also talked about how -

OpenAI Prepares Codex Upgrade with GPT-5.2-Codex-Max
By
–


BREAKING : OpenAI is preparing to upgrade Codex, likely with GPT-5.2-Codex-Max. One last leap this year? Caribou
-

Small Model Outperforms DeepSeek-r1 With Recursive Reasoning
By
–
Recursive reasoning beats multi-billion-parameter models You can now easily train your own 7M param model from scratch and outperform DeepSeek-r1 on ARC-AGI 1 We provide a simple speedrun script that handles setup, training, and eval in one go.
-
Gemini 3 Flash Demonstrates Frontier Reasoning and Multimodal Capabilities
By
–
In this demo, you’ll see Gemini 3 Flash’s frontier-level reasoning and multimodal capabilities on display. The model is able to simultaneously conduct complex geometric calculations while processing complex inputs (video and image).
— Google AI (@GoogleAI) 17 décembre 2025
You can play around with the slingshot in… pic.twitter.com/vcleTTRzbYIn this demo, you’ll see Gemini 3 Flash’s frontier-level reasoning and multimodal capabilities on display. The model is able to simultaneously conduct complex geometric calculations while processing complex inputs (video and image). You can play around with the slingshot in
-
Gemini 3.0 Flash: Near Gemini Pro Performance with Faster Speed
By
–
It's not the end of the year yet, amazing how ⚡⚡⚡ the progress is, for example Humanity Last Exam, almost as good as Gemini 3.0 Pro but much faster! Read more https://t.co/97MH68nlQc https://t.co/zTT0PFbZq7 pic.twitter.com/ZJqsNCZ7xt
— Thang Luong (@lmthang) 17 décembre 2025It's not the end of the year yet, amazing how the progress is, for example Humanity Last Exam, almost as good as Gemini 3.0 Pro but much faster! Read more https://
blog.google/products/gemin
i/gemini-3-flash
… -
Gemini 3 Flash delivers rapid AI model performance updates
By
–
Delivered with the speed of ! Thanks for sharing and keep it coming, we'd love to see anything you build with Gemini 3 Flash.
-

LangSmith Tracing with Claude Code Agents Feedback Loop
By
–
LangSmith + Claude Code / Deepagents Pairing LangSmith tracing w/ code agents provides a powerful feedback loop. Here, we show examples of that w/ langsmith-fetch + Claude Code / Deepagents. langsmith-fetch CLI: https://
github.com/langchain-ai/l
angsmith-fetch
… : https://
youtu.be/zpgFl4N4DIc -
GPT 1.5 Image Available, Gemini 3.0 Flash Coming Soon
By
–
– GPT 1.5 Image is LIVE on ChatLLM
– Gemini 3.0 Flash will be live in a couple of hours