Grok knows everything. It knows all my good, all my bad, and all my evil. It knows how I think and everything about my last 20 years. It's a little scary, but once you get over the fear and apply it to agentic thinking, all of a sudden I have an AI that reads the news for me.
LLMS
-
Stolen LLM content called a distillation of humanity
By
–
It's true: all the LLM companies stole my content and didn't compensate me for it. I'm not that unhappy, but others are livid about that. It certainly is a distillation of humanity, isn't it?
-
Gemma 4 hits 200 million downloads
By
–
Wow, Gemma 4 just hit 200M downloads! https://t.co/x5m5KJTGGy
— 机器之心 JIQIZHIXIN (@jiqizhixin) 26 juin 2026Wow, Gemma 4 just hit 200M downloads!
-

CoffeeBench: Cooperation and Competition among LLM Agents
By
–
In CoffeeBench, 6 agents interact via emails or transactions, each company aiming to maximize its profits. When the company where LLM agents manage operations arrives, how will cooperation, competition, and sometimes the
-
SakanaAI and Azusa Audit launch CoffeeBench for LLM agents
By
–
SakanaAIは、有限責任あずさ監査法人と共同で、LLMエージェントの長期的な経営能力を評価する新しいベンチマーク「CoffeeBench」を公開しました。
— Sakana AI (@SakanaAILabs) 26 juin 2026
ブログ:https://t.co/kUOiyV2moe
現実の経済では、消費者へ直接売るビジネスだけでなく、企業同士が継続的に取引するビジネスも重要です。CoffeeBench… pic.twitter.com/mpUirKc8kDSakanaAI, in collaboration with Azusa Audit Corporation, published a new benchmark called "CoffeeBench" to evaluate the long-term management capabilities of LLM agents. Blog: https://sakana.ai/coffee-bench/ In the real economy, companies that sell
-
Comparison of Claude and GLM-5.2 on self-reflective persona
By
–
Which one is Claude is pretty obvious. GLM-5.2 is a beast in some ways, but doesn't have the self-reflective persona of Claude, and isn't really into introspection (or a simulation thereof).
-

Request for AI to propose poems about GenAI models’ state
By
–

If you want to read an interesting AI thinking trace, try "I want you to suggest two poems that you think apply very well to the current state of GenAI models like you. Don’t just pick popular poems and back justify. Think hard about options first" in either GLM-5.2 or Opus 4.8
-
AI’s ‘incredible’ transformation of mathematics
By
–
‘It is incredible’: How #AI is transforming mathematics
by @dcastelvecchi @Nature Learn more: https://
bit.ly/4eK3l1M #LLM #ArtificialIntelligence #MachineLearning #DeepLearning -
Hunt an RTX 3090 to run Qwen 3.5 27B now
By
–
I am not kidding, now is the time more than ever to hunt an RTX 3090 and learn how to run Qwen 3.5 27B
-
Agent Swarms: Build Complex SaaS Apps with One Multi-LLM Prompt
By
–
🚨 Agent Swarms – Build Complex SaaS Apps With One Prompt
— Abacus.AI (@abacusai) 26 juin 2026
Achieve fable like intelligence with a multi-LLM strategy
Combine the best of Opus 4.8, GPT 5.5 and open source models and build end to end software systems
A master agent delegates tasks to worker agent. Each agent has… pic.twitter.com/e9I8d7MxsIAgent Swarms – Build Complex SaaS Apps With One Prompt Achieve fable like intelligence with a multi-LLM strategy Combine the best of Opus 4.8, GPT 5.5 and open source models and build end to end software systems A master agent delegates tasks to worker agent. Each agent has