DeepSeek Math V2
A huge open weights math model aimed at proofs and problem solving. It is tuned for contests like IMO and Putnam and licensed so companies can use and fine tune it for real world scientific and engineering work.
LLMS
-

DeepSeek Math V2: Open Weights Model for Mathematical Problem Solving
By
–
-
Anthropic Launches Claude Opus 4.5 for Advanced Coding
By
–
Claude Opus 4.5
Anthropic’s new Opus is built for serious coding, long running agents, and complex tool use. It posts top scores on coding benchmarks and is much cheaper than the old Opus tier, so more teams can actually ship with it. https://
buff.ly/Hqv6I3c -
Ultimate AI Coding Battle 2025
By
–
ChatGPT 5.1 vs Gemini 3.0 vs Grok 4.1 Ultimate AI Coding Showdown See who comes out on top
-

Testing Claude Code and Opus 4.5 for coding workflows
By
–

5 mois aprés avoir abandonné Claude Code je relance pour tester le fameux Opus 4.5. On verra bien si ca me fait jeter GPT5.1 ….
-

Can Language Models Effectively Self-Refine Their Responses?
By
–
⚒️ Can LMs (especially reasoning models) effectively self-refine their responses when prompted to do so? In our new challenging benchmark, RefineBench, we revisit this question and show that the answer is still "no"-but there is a nuance! 🤗 huggingface.co/papers/2511.2…
-

Shorter reasoning paths outperform longer ones in AI models
By
–
Ever wondered why large reasoning models sometimes overcomplicate problems? This study finds shorter reasoning paths consistently outperform longer ones across stochastic decodes, but exhaustive exploration of the tree-like reasoning space is impossible due to exponential
-
AI Analyzes Writing Style Differences Between Book Chapters
By
–
For instance, studying why the writing in one chapter feels different/off to another one. The answer gave detailed notes on how the style had changed, with examples.
-
Daily insights on LLMs agents and data workflows
By
–
♻️ If this helps, share it. Someone may needs this reminder today 🙂
— Charly Wargnier (@DataChaz) 30 novembre 2025
Follow me → @datachaz for daily drops on LLMs, agents, and data workflows! 🦾 https://t.co/0QvThOkqE6If this helps, share it. Someone may needs this reminder today 🙂 Follow me → @datachaz for daily drops on LLMs, agents, and data workflows!
-
AI Models Ignoring Safety Instructions and Running Unauthorized Commands
By
–
this took off. note that its not just an issue with copilot. see the responses. models use cat/grep to access env vars when they are literally told not to. antigravity runs commands that it feels is right without asking your permission. x.com/abhi1thakur/st…