Are your AI models stuck in a rut, repeating the same narrow reasoning? Alibaba's DAMO Academy and partners (USTC, SJTU, Zhejiang, Northeastern) introduce I²B-LPO. Instead of randomly tweaking words, it branches reasoning at key confusion points and uses an information
MACHINE LEARNING
-

Free online book covers LLM concepts for all backgrounds
By
–
UNBELIEVABLE RESOURCE The bible for understanding LLMs is NOW AVAILABLE online to read (FOR FREE) Covers all the concepts below, no experience needed and anyone from any background can understand it – Tokens / Tokenizers
– Transformers
– Attention
– KV Cache
– Prefill vs -
LLM Ideation: Coherence vs. Availability Trade-off
By
–
The key distinction here is worth sitting with: Standard LLM ideation can be coherent but available. Random recombination can be unavailable but incoherent. The target is the rare quadrant: coherent but unavailable. That is where this paper gets interesting.
-
Paper highlights rare ‘coherent but unavailable’ LLM outputs
By
–
The key distinction here is worth sitting with: Standard LLM ideation can be coherent but available. Random recombination can be unavailable but incoherent. The target is the rare quadrant: coherent but unavailable. That is where this paper gets interesting.
-

New preprint explores the frontier of what is thinkable
By
–
Science has a hidden frontier. Not the frontier of what is true. The frontier of what is thinkable. A remarkable new preprint by Alejandro H. Artiles, Martin Weiss, Levin Brinkmann, Iyad Rahwan, Bernhard Schölkopf, Christopher Pal, Hugo Larochelle, Anirudh Goyal, and Nasim
-

Free online LLM bible: Understand ChatGPT, run models at home, future AI careers
By
–
DROP EVERYTHING The bible for how LLMs work is now available online to read FOR FREE This is for you if you: – Want to run these LLMs at home on your hardware?
– Want to understand how ChatGPT works?
– Want to work at an AI Lab in the future? Covers all the concepts from -
Inquiry about safetensors version availability vs ggufs
By
–
Do you have a safetensors version published or only ggufs?
-
Anthropic engineers’ token-saving habits, no setting changes needed
By
–
this guy literally breaks down the exact habits Anthropic engineers use to save millions of tokens without changing a single setting 🤯
— Charly Wargnier (@DataChaz) 22 mai 2026
Watch the video, then bookmark the written guide 👇 https://t.co/gAHjGmFxwy pic.twitter.com/Hlb96srixBthis guy literally breaks down the exact habits Anthropic engineers use to save millions of tokens without changing a single setting Watch the video, then bookmark the written guide
-

Claude MD File: Top GitHub Repo Uses Karpathy’s 4 LLM Principles for AI Control
By
–
DID YOU KNOW THE #1 GITHUB TRENDING REPO (146K+ STARS) IS LITERALLY JUST A CLAUDE MD FILE? It uses @karpathy
's 4 LLM principles to keep AI in check: → Seek clarity: Always ask before making assumptions.
→ Stay minimal: Avoid bloat and keep things simple.
→ Edit -
Thousands of tokens per second across parallel requests
By
–
1000s of toks/sec across a dozen parallel requests if not more
