What if AI could scale its reasoning without wasting compute? Researchers from NUS, Georgia Tech, and other institutions present PRISM — a test-time scaling method for discrete diffusion language models (dLLMs). It uses hierarchical search to prune and reallocate compute
