Let Agents Design Agents Memento-Skills is a self-evolving agent framework where agents learn from failures and rewrite their own skills. Most agent frameworks treat skills as static. You write them once, load them into context, and hope they work. When they fail, you debug
LLMS
-
Open Source Sustainability Challenges with Large AI Models
By
–
Open source project are usually subsidized by other income and revenue streams (e.g., developers' corporate job and other side projects, companies' products) to keep it sustainable. With open weight models, it's an increasingly large price tag (because of GPU demands), and there
-
Claude Cowork: Desktop Agent for Autonomous Job Application Filling
By
–
QU'EST-CE QUE CLAUDE COWORK ? Claude Cowork = App desktop qui donne à Claude son propre navigateur Google Chrome. PAS un chatbot. C'EST un agent qui : → Ouvre onglets → Lit offres d'emploi → Remplit candidatures → Prend décisions Basé sur règles que vous définissez
-

Claude’s Global Adoption: USA, Europe Lead While China Lags
By
–
Claude's use in the USA is no surprise. But apparently, Europe has also developed above-average use of Claude. China lags far behind. At least officially, there is no use of Claude (although we know that distillation was a significant part of its production). Funnily enough:
-

DeepSeek Releases Larger Base Model Amid Training Silence
By
–
"A new, much larger (DeepSeek) base model will be released soon", from DeepSeek staff. I'm currently wondering why there's been so much silence surrounding DeepSeek. The last report stated that they attempted to train on Huawei chips but failed ("DeepSeek AI model failed to
-
Best Model for Actual Machine Learning Work
By
–
For actual ML work it's still the best overall model around.
-
Google’s TurboQuant Algorithm Enables Local Execution of Large LLMs
By
–
Esta es potencialmente la noticia más importante del año.
— SONIA (@S0N_IA_) 25 mars 2026
Google acaba de lanzar TurboQuant. Un algoritmo que hace que los modelos LLM sean más pequeños y rápidos, sin perder calidad.
Esto significa que ahora un Mac Mini de 16 GB puede ejecutar modelos de IA INCREÍBLES.… https://t.co/O271fYGWxUThis is potentially the most important news of the year. Google has just launched TurboQuant. An algorithm that makes LLM models smaller and faster, without losing quality. This means that now a 16 GB Mac Mini can run INCREDIBLE AI models. Completely local, free, and secure.
-
Q-Priming Improves Agent Reliability Through Clarification Requests
By
–
The Q-priming part is really interesting. Making models 5x more likely to ask for clarification instead of guessing wrong is exactly what we need for agentic workflows. Right now most agents just confidently charge ahead with bad assumptions.
-
Distilling Opus reasoning into 27B local model for production
By
–
Distilling Opus reasoning into a 27B model that runs locally is incredible. The fact that tool calling works too makes this actually useful for production, not just benchmarks. Super excited to test its limits 😀
-

Google DeepMind TurboQuant Redefines AI Model Compression Efficiency
By
–
Absolute GAME CHANGER for local AI @GoogleDeepMind
’s new TurboQuant algorithm achieves extreme AI compression, completely redefining model efficiency. Usually, shrinking an AI makes it “dumber.” Not this time. It’s an algorithm that makes LLMs smaller and faster without
