Another cute comic for this work: a nice collaborative story between #Aletheia and human experts for cracking Erdős problems, thanks @AwaitSeedIgnite! I wish we have ComicBanana for all papers 🙂 First four pages here and full one in this link https://
drive.google.com/file/d/1XecW9W
tDDNl-XE5-lm9DB71KX0FcDMH5/view?usp=sharing
….
AGENTS
-
AI Collaborative Research on Erdős Problems with Comic Documentation
By
–
-
Urging focus on agentic capabilities for LLM breakthrough
By
–
But please @GuillaumeLample first you need to focus on developing more and better agentic capabilities of your model. If you can achieve a sota and huge breakthrough on a LLM benchmark agentic capabilities with a SML… that can change the whole industrie !!! That can be so
-

New AGI Benchmark Achievement with Multi-Model Systems Outperforming Single Models
By
–
A new 72% acheivement submission for ARC-AGI-2. So far, it is the second multi-model system that outperformed single-model solutions. "It runs the same task through GPT-5.2, Gemini-3, and Claude Opus 4.5 in parallel." We need new benchmarks
-

AI Agent Proliferation: Risks and Control in Organizations
By
–
AI agents are popping up everywhere, and fast! 🤯 Do you know how many are running in your organization? Hear why #AI agent proliferation is so scary and what happens when things go wrong in this interview with @RubrikInc GM of AI Dev Rishi 👉 go.rbrk.co/xobwcp
→ View original post on X — @predibase, 2026-02-03 18:09 UTC
-

New Book on Agentic Architectural Patterns and Multi-Agent Systems
By
–
New release from @PacktDataML @PacktPublishing "Agentic Architectural Patterns for Building Multi-Agent Systems: Proven design patterns and practices for GenAI, agents, RAG, LLMOps, and enterprise-scale AI systems" See it at https://
amzn.to/3MaHy8T 𝕋𝕒𝕓𝕝𝕖 𝕠𝕗 -
AI Transforms Customer Support in 90 Days at CoreWeave
By
–
Adding AI to an existing workflow doesn’t need to be a lengthy, disruptive process. We helped @CoreWeave transform its customer support efforts in 90 days with our agentic platform, North. Read the full story to see what they achieved.
-
Boring AI: Agents for Operational Work and Business Outcomes
By
–
@rohanpaul_ai We totally agree. The real value of AI is in quietly taking on the boring, operational work that keeps businesses running. That’s exactly what we’re building with Boring AI: agents designed for accuracy, reliability and outcomes, not hype. Check out our work:
-
Exploring Multi-Agent Ecology and Shared Knowledge Scratchpads
By
–
Absolute bars: “Sites like moltbook function as a giant, shared, read/write scratchpad for an ecology of AI agents – how might these agents begin to use this scratchpad to a) influence future ‘blank slate’ agents arriving at it the first time, and b) unlock large-scale
-
Building Trust in Autonomous AI: Governance Blueprint
By
–
Building trust in autonomous AI: A governance blueprint for the agentic era
#AI #AIio #AIInnovation #ML #DataScience #Futureofwork @Scobleizer @AndrewYNg @drfeifei @KirkDBorne @fchollet @rowancheung @antgrasso -

Kaggle and Google DeepMind Introduce New Benchmarks for AI Agentic Capabilities
By
–
In partnership with @GoogleDeepMind
, @Kaggle announced today an expansion of the #Kaggle Game Arena, introducing new benchmarks designed to test critical #AI skills required for real-world agentic capabilities. Key updates include:
New werewolf benchmark
New poker benchmark