"ablation studies are not visible to the user" 😛
MACHINE LEARNING
-

A model more capable than a year ago, priced below Opus 4.1
By
–
If you really look at it, this model, much more capable than what we had a year ago, is released with a price below Opus 4.1.
-
The importance of self-verification loops in the era of powerful models
By
–
We talk a lot about how important it is to set up self-verification loops. Especially in the age of powerful models that can run for long periods of time, self-verification is a key ingredient that enables the model to run for much longer, delivering a result that is closer to… https://t.co/NHiral0F9j
— Boris Cherny (@bcherny) 9 juin 2026We talk a lot about the importance of setting up self-verification loops. Especially in the era of powerful models that can operate for long periods, self-verification is a key ingredient that allows the model to function much more
-

xAI and Gopuff build personalized shopping assistant with chat, voice, image models
By
–
Learn more about our work with @gopuff to build a personalized shopping assistant with chat, voice, and image models https://
x.ai/news/grok-gopu
ff
… -
Limiting Claude for cutting-edge LLM development widens the gap
By
–
“limiting Claude’s effectiveness for requests aimed at developing cutting-edge LLMs” And that’s how they plan to widen the gap even further
-

Anthropic’s new Fable 5 safeguards quietly limit effectiveness
By
–
Anthropic’s new Fable 5 safeguards are fascinating. When the model is used for frontier LLM development, it apparently does not simply refuse or warn the user. Instead, it quietly limits its own effectiveness through techniques like prompt modification, steering vectors, and
-
Mayo Clinic AI Detects Pancreatic Cancer 3 Years Early
By
–
Mayo Clinic researchers have developed an AI model that detects pancreatic cancer on routine CT scans up to three years before clinical diagnosis.
— The Rundown AI (@TheRundownAI) 9 juin 2026
Published in Gut, the study tested the model, called REDMOD, on nearly 2,000 scans, including pre-diagnostic scans originally read… pic.twitter.com/FLcPuuus1PResearchers at Mayo Clinic have developed an AI model that detects pancreatic cancer on routine CT scans up to three years before clinical diagnosis. Published in Gut, the study tested the model, called REDMOD, on nearly 2,000 CT scans.
-

Apple’s Core AI runs models entirely on-device
By
–
Apple finally did it. Its new framework, Core AI, runs models entirely on Apple silicon, so inference happens on the user's device with zero server calls and zero token bills. That means Qwen, Mistral, and SAM3 running natively across iPhone, iPad, Mac, and Vision Pro. It's a
-

Anthropic would limit capabilities to maintain competitive advantage
By
–
Pretty crazy this that is being shared where Anthropic would be limiting the model's capabilities when used to improve and create better LLMs. They sell it as a security measure but it is clear that they do it to maintain their competitive advantage.
-

OSCAR: 2-bit KV cache for LLMs without accuracy loss
By
–
Can LLMs run on ultra-low-bit memory without tanking accuracy? Researchers from Together AI, University of Sydney, and UIUC present OSCAR — a method that uses offline, attention-aware covariance analysis to design fixed rotations and clipping thresholds for 2-bit KV cache