i dont think it was darios fault in the long run. i think at any point where AI becomes too powerful it would have been regulated. that being said, you cant ban open source that easily.
AI
-
Securely monitor long-running agents from your phone
By
–
4/ The cherry on top!
— Charly Wargnier (@DataChaz) 26 juin 2026
You can securely monitor long-running agents from your phone (works on both iOS and Android)
.. while the session keeps running locally on your machine 🔥 pic.twitter.com/k1ybWW1ntK4/ The cherry on top! You can securely monitor long-running agents from your phone (works on both iOS and Android) .. while the session keeps running locally on your machine
-
RMUX: True multiplexer for agent-based workflows with dynamic panes
By
–
3/ because RMUX is a true multiplexer first, you aren't just launching a fragile web wrapper. You get a real environment perfectly designed for modern, agent-based workflows. → Safely inspect long-running AI shells
→ Resize splits and manage panes dynamically
→ Share -

Loop Engineering: Building Systems That Prompt Agents Instead
By
–

A SENIOR ANTHROPIC ENGINEER JUST DROPPED AN 11-PAGE PDF ON LOOP ENGINEERING. The core shift: stop prompting the agent. Build the system that prompts it. Inside the autonomous loop: – Discover → Finds its own work (failing CI, open issues).
– Isolate → Uses separate git -
Gemma 4 hits 200 million downloads
By
–
Wow, Gemma 4 just hit 200M downloads! https://t.co/x5m5KJTGGy
— 机器之心 JIQIZHIXIN (@jiqizhixin) 26 juin 2026Wow, Gemma 4 just hit 200M downloads!
-

CoffeeBench: Cooperation and Competition among LLM Agents
By
–
In CoffeeBench, 6 agents interact via emails or transactions, each company aiming to maximize its profits. When the company where LLM agents manage operations arrives, how will cooperation, competition, and sometimes the
-
SakanaAI and Azusa Audit launch CoffeeBench for LLM agents
By
–
SakanaAIは、有限責任あずさ監査法人と共同で、LLMエージェントの長期的な経営能力を評価する新しいベンチマーク「CoffeeBench」を公開しました。
— Sakana AI (@SakanaAILabs) 26 juin 2026
ブログ:https://t.co/kUOiyV2moe
現実の経済では、消費者へ直接売るビジネスだけでなく、企業同士が継続的に取引するビジネスも重要です。CoffeeBench… pic.twitter.com/mpUirKc8kDSakanaAI, in collaboration with Azusa Audit Corporation, published a new benchmark called "CoffeeBench" to evaluate the long-term management capabilities of LLM agents. Blog: https://sakana.ai/coffee-bench/ In the real economy, companies that sell
-
Comparison of Claude and GLM-5.2 on self-reflective persona
By
–
Which one is Claude is pretty obvious. GLM-5.2 is a beast in some ways, but doesn't have the self-reflective persona of Claude, and isn't really into introspection (or a simulation thereof).
-

Request for AI to propose poems about GenAI models’ state
By
–

If you want to read an interesting AI thinking trace, try "I want you to suggest two poems that you think apply very well to the current state of GenAI models like you. Don’t just pick popular poems and back justify. Think hard about options first" in either GLM-5.2 or Opus 4.8
-
AI’s ‘incredible’ transformation of mathematics
By
–
‘It is incredible’: How #AI is transforming mathematics
by @dcastelvecchi @Nature Learn more: https://
bit.ly/4eK3l1M #LLM #ArtificialIntelligence #MachineLearning #DeepLearning