the way i see it, the last twelve months of AI research can be summed up in just two big breakthroughs: [i] reasoning ('test-time compute') – new ways to train models that can use more tokens to generate better answers. they mostly rely on RL with verifiable rewards [ii]
Two Major AI Breakthroughs: Reasoning and Test-Time Compute
By
–
