Diffusion LLMs have been losing to autoregressive models for two years. UC San Diego just found the one thing they're actually better at: not generating, drafting. DFlash hits 6x lossless acceleration on Qwen3-8B. 2.5x faster than EAGLE-3. Check out the detailed breakdown
DFlash 6x lossless acceleration on Qwen3-8B
By
–
