BullshitBench: sorry to say but DeepSeek v4 did really badly, towards the bottom of the table, whether it is high or low reasoning.
RESEARCH
-

AI Models Learn Self-Improvement Without External Rewards
By
–
Can an AI teach itself to reason better without any outside reward? Researchers from CUHK, Shenzhen, SJTU, and CUHK present SePT. They let a language model generate its own reasoning examples by using "low-temperature" (more focused) responses, then train on that new data in a
-

Abstract Chain-of-Thought: Efficient Latent Reasoning Without Words
By
–
"Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought" Do reasoning models really need to think in words? This paper replaces long verbal CoT with a short learned sequence of abstract tokens that acts like a latent scratchpad. Warmed up from
-

Scientific Theory of Deep Learning Emerges
By
–
“There Will Be a Scientific Theory of Deep Learning” This paper argues that a real scientific theory of deep learning is beginning to emerge. Not a theory that tracks every neuron individually, but a physics-like theory of learning itself. One that aims to characterize how
-
Research on Prompt Compression Using Draft Models Accepted to ICLR
By
–
Another research accepted to ICLR 2026 We explored a new way to shrink long prompts using smaller draft models from different model families, no retraining needed. Faster time to first token, with performance holding strong. Take a look @UrmishThakker
-

Bloom: Stanford’s LLM-Based AI Health Coaching App
By
–
Most health apps prescribe rigid plans. Bloom, an LLM-based health coaching app created by @Stanford researchers, applies evidence-based coaching methods. A promising demonstration of human-centered AI: https://
hai.stanford.edu/news/an-ai-hea
lth-coach-could-change-your-mindset
… -
Incentive Misalignment: Quantity Over Quality in AI Systems
By
–
The problem is that the incentives push for "more" over "better" Paper: https://
pubsonline.informs.org/doi/full/10.12
87/orsc.2026.ed.v37.n3
… -

AI in Science: Quality Over Quantity in Research Systems
By
–
Very cool analysis of the submissions to a major management journal that shows how much the system of science, built for humans, is under strain as a result of AI. AI can be used to do better science or it can be used to just do more stuff. The danger is that "more" is winning
-

Accidental Medical Discoveries: Serendipity in Healthcare Innovation
By
–
Big innovations that were discovered by mistake!
Nice summary by @Zlatimeyer Can add coronary angiography by Mason Sones to the list, and many others in medicine.
Gift link https://
wsj.com/business/us-in
ventions-mistake-discovery-8d0ff716?st=WkraPi&reflink=desktopwebshare_permalink
… -
Precise Video Language Model with Human-AI Oversight Framework
By
–
Building a Precise Video Language with Human-AI Oversight
— AK (@_akhaliq) 27 avril 2026
paper: https://t.co/przLkLaUhH pic.twitter.com/yXANHZ8J31Building a Precise Language with Human-AI Oversight paper: https://
huggingface.co/papers/2604.21
718
…
