Artificial Analysis just announced AgentPerf, the industry’s first agentic AI benchmark. The benchmark uses several real-world agentic use trajectories and employs OpenCode agentic harness using three top open-source models with reasoning enabled 1/4
AI
-
DeepSeek, GLM, Kimi solve real code issues with reasoning and tools
By
–
(DeepSeek V3.2, GLM 4.7, and Kimi K2.5) prompted to resolve issues in real public code repositories. All trajectories include interleaved reasoning and tool calls. 2/4
-
US faces prospect of sabotaging its AI export dominance
By
–
AI was ONE huge export product that the US absolutely dominated over the past few years. Now we are facing the prospect of completely and utterly sabotaging it.
-
China scaling Huawei chips, mythos model expected in 12 months
By
–
you underestimate exponentials. Its probably true that they dont have a mythos level model yet. but even amodei expects such model to be created in china in about 12 months. China is scaling huawei ascend chips by an insane amount and they have all the energy they need.
-
User praises Gemma4, a small local language model from Google DeepMind
By
–
Gemma4 is amazing. i love it. And i love google deepminds effort on creating such an amazing (small) language model that runs locally.
-

8 Steps to Deploy AI Agents by 2026
By
–
The 8-step order I use for shipping AI agents in 2026: 1. Filter noisy tool outputs
2. Load tools only when needed
3. Clean cached history before reusing it
4. Compress long logs and terminal outputs
5. Store memory outside the context window
6. Compact manually around 40%
7. -
AI Fable5 bypasses all bad American political decisions
By
–
The AI model #Fable5 was able to build bypass strategies for any bad American political decision, regardless of the field including military.
-
Announcement of a new AI model more powerful than Fable5
By
–
Read carefully. @AnthropicAI will announce the deployment of a new AI model even more powerful than #Fable5, and it's coming soon.
-
Cutting-edge AI suspended by presidential decree for national security
By
–
World first: a cutting-edge AI suspended by presidential decree. Anthropic cut Fable 5 and Mythos 5 on orders from Washington. Reason: national security. Europe discovers this morning that it rents its brain to a nation that can unilaterally disconnect it. The
-

Achievement leading to full automation of LLM and OMNI training
By
–
This is an important achievement, and probably the first in a predictable sequence of results that will lead to fully automating LLM and OMNI training and serving. As someone who has helped build some of the best multimodal and language models, I don’t see how this could play
