Giving your models more time to think before prediction, like via smart decoding, chain-of-thoughts reasoning, latent thoughts, etc, turns out to be quite effective for unblocking the next level of intelligence. New post is here 🙂 “Why we think”: lilianweng.github.io/posts/2…
→ View original post on X — @lilianweng, 2025-05-17 15:09 UTC