A new 72% acheivement submission for ARC-AGI-2. So far, it is the second multi-model system that outperformed single-model solutions. "It runs the same task through GPT-5.2, Gemini-3, and Claude Opus 4.5 in parallel." We need new benchmarks
LLMS
-

New Book on Agentic Architectural Patterns and Multi-Agent Systems
By
–
New release from @PacktDataML @PacktPublishing "Agentic Architectural Patterns for Building Multi-Agent Systems: Proven design patterns and practices for GenAI, agents, RAG, LLMOps, and enterprise-scale AI systems" See it at https://
amzn.to/3MaHy8T 𝕋𝕒𝕓𝕝𝕖 𝕠𝕗 -
User Learns English in 5 Weeks with ChatGPT Prompts, Outperforms Duolingo
By
–
R.I.P. DUOLINGO. 3 ans à apprendre l’ANGLAIS. Rien n’a marché. ChatGPT l’a fait en 5 semaines. Voici les 6 prompts que j’ai utilisés: [ Ajoutez en signetpour ne pas perdre ! ]
-

Technical approaches to AI model context window optimization
By
–
I do think that better compaction and teaching the models to re-learn context post compaction if they are unsure solves the need for really long context windows to an extent. That said, GPT 5.2 and the codex variants both support 400K context window which is a step up.
-
Introduction of LingBot-World: A New AI World Model
By
–
4/
LingBot-World, a world model that achieves industry-leading performance in video quality, dynamic fidelity, long-term consistency, and interactivity – learn about it here: -
LingBot-VLA: A Vision-Language-Action Model for Robotics
By
–
3/
LingBot-VLA, a vision-language-action model designed to serve as a “universal brain” for real-world robotics – learn about it here: https://
afp.com/en/infos/robby
ant-open-sources-lingbot-vla-universal-brain-robots
… -
Jamba2: Open Source Model for February On-Device Applications
By
–
Model to try out in February: Jamba2. Fully open source under Apache 2.0, includes 3B (dense) model for on-device apps, intended for grounded QA workflows that don't call for the heavy “thinking token” overhead of reasoning models. Further details:
-
Jamba2: Open Source Model for On-Device QA Applications
By
–
Model to try out in February: Jamba2. Fully open source under Apache 2.0, includes 3B (dense) model for on-device apps, intended for grounded QA workflows that don't call for the heavy “thinking token” overhead of reasoning models. Further details:
-
Play-testing the Genie 3 AI model
By
–
My pleasure! Dropped a new video with a lot more genie 3 play testing:
-
Exploring Multi-Agent Ecology and Shared Knowledge Scratchpads
By
–
Absolute bars: “Sites like moltbook function as a giant, shared, read/write scratchpad for an ecology of AI agents – how might these agents begin to use this scratchpad to a) influence future ‘blank slate’ agents arriving at it the first time, and b) unlock large-scale
