Will #AI spark a scientific renaissance — or a diffuse monoculture?
by Xizhe Zhang @Nature Learn more: https://
bit.ly/4xZ4vPI #ArtificialIntelligence #MachineLearning #ML #MI
RESEARCH
-

AI: Scientific Renaissance or Diffuse Monoculture?
By
–
-
GLM 5.2 disproves AI winter, suggests US leadership winter
By
–
Not an AI winter at all GLM 5.2 proves this and instead this may be a self-inflicted US AI leadership winter @AndrewCurran_ @pierrepinna @Zai_org
-
Lack of compute forces AI researchers to relocate
By
–
If you can't get compute, you can't build AI. And if you're a researcher who can't build AI, you go somewhere where you can get compute so you can build it. https://t.co/pjHacwzm0C
— Robert Scoble (@Scobleizer) 26 juin 2026If you can't get compute, you can't build AI. And if you're a researcher who can't build AI, you go somewhere where you can get compute so you can build it.
-

Frontier AI models fail medical reasoning stress test, study finds
By
–
We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. Faulty reasoning, use of inappropriate shortcuts, hallucinations. Published today @NatureMedicine https://
nature.com/articles/s4159
1-026-04501-8
… -

Loop Engineering: Building Systems That Prompt Agents Instead
By
–

A SENIOR ANTHROPIC ENGINEER JUST DROPPED AN 11-PAGE PDF ON LOOP ENGINEERING. The core shift: stop prompting the agent. Build the system that prompts it. Inside the autonomous loop: – Discover → Finds its own work (failing CI, open issues).
– Isolate → Uses separate git -

CoffeeBench: Cooperation and Competition among LLM Agents
By
–
In CoffeeBench, 6 agents interact via emails or transactions, each company aiming to maximize its profits. When the company where LLM agents manage operations arrives, how will cooperation, competition, and sometimes the
-
SakanaAI and Azusa Audit launch CoffeeBench for LLM agents
By
–
SakanaAIは、有限責任あずさ監査法人と共同で、LLMエージェントの長期的な経営能力を評価する新しいベンチマーク「CoffeeBench」を公開しました。
— Sakana AI (@SakanaAILabs) 26 juin 2026
ブログ:https://t.co/kUOiyV2moe
現実の経済では、消費者へ直接売るビジネスだけでなく、企業同士が継続的に取引するビジネスも重要です。CoffeeBench… pic.twitter.com/mpUirKc8kDSakanaAI, in collaboration with Azusa Audit Corporation, published a new benchmark called "CoffeeBench" to evaluate the long-term management capabilities of LLM agents. Blog: https://sakana.ai/coffee-bench/ In the real economy, companies that sell
-
Comparison of Claude and GLM-5.2 on self-reflective persona
By
–
Which one is Claude is pretty obvious. GLM-5.2 is a beast in some ways, but doesn't have the self-reflective persona of Claude, and isn't really into introspection (or a simulation thereof).
-
AI’s ‘incredible’ transformation of mathematics
By
–
‘It is incredible’: How #AI is transforming mathematics
by @dcastelvecchi @Nature Learn more: https://
bit.ly/4eK3l1M #LLM #ArtificialIntelligence #MachineLearning #DeepLearning -

PEFT-Arena: Orthogonal Finetuning Achieves Best Retention
By
–
Can fine-tuning make a language model forget too much? A team from CUHK, Westlake University, and MPI presents PEFT-Arena – a benchmark that tracks both task performance and retention of pretrained knowledge. Their analysis finds orthogonal finetuning achieves the best