8). Z1 Z1 is a new method for making LLMs more compute-efficient at test time, especially during reasoning. It train LLMs with short and long code-based reasoning trajectories, and then dynamically adjusts reasoning depth during inference.
LLMS
-

Why LLMs Focus Attention on First Token
By
–
5). Why do LLMs Attend to First Token? This new paper explains why LLMs obsessively focus attention on the first token — a phenomenon known as an attention sink.
-

MedAgentSim: Automated LLM Hospital Simulation for Doctor-Patient Interactions
By
–
6). MedAgentSim MedAgentSim is a fully automated, open-source hospital simulation where LLM-powered agents simulate doctor-patient interactions in dynamic diagnostic settings.
-

RARE: New Reasoning-Focused LLM Training Paradigm
By
–
4). Retrieval-Augmented Reasoning Model Introduces RARE, a new paradigm for training domain-specific LLMs that focuses on reasoning, not memorization.
-

Cohere Launches Command A: 111B Parameter Enterprise LLM
By
–
2). Command A: An Enterprise-Ready LLM Cohere announced Command A, a 111B parameter open-weights LLM built for enterprise-grade RAG, agents, code, and multilingual tasks.
-
Red Hat SVP Misuses Open Source Label for Meta Llama
By
–
When an SVP at Red Hat unironically calls Llama 4 "open source" (despite the same 700M user limit, etc. @OpenSourceOrg has criticized: https://
opensource.org/blog/metas-lla
ma-license-is-still-not-open-source
…), it raises the question as to what damage AI has done to traditional views on "open source." -
10M Token Processing and Infinite Context Length Breakthroughs
By
–
Wonder what compute times we will see with 10M input tokens on a rather beefy system. The NiH tests are also insane, It’s hard to believe we are moving towards infinite token context lengths with near perfect retrieval.
-

DeepSite: an ultra-high-performance AI that is 100% free surpasses Bolt and Manus
By
–
Forget Bolt, Manus AI and the others: here is DeepSite, an ultra-high-performance AI that is 100% free! The complete tutorial → https://youtu.be/l13s3lVNK4w #AITools #ArtificialIntelligence #Deepsite #Deepseek #DeepSeekv3
-
Gemma 3 Open Source Models Run Single GPU TPU
By
–
Fwiw, this exact reason is why we made the Gemma 3 open source models something that developers could easily run on a single GPU or TPU.
-
Meta Releases Open-Weight Model with 10M Token Context
By
–
I don't look at a computer for a few hours today and come back to find that Meta dropped an open-weight model with a 10 MILLION TOKEN Context Window! Holy crap!