Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding
Paper: https://
arxiv.org/pdf/2505.22618
.pdf
…
Code: https://
github.com/NVlabs/Fast-dL
LM
…
Project Page: https://
nvlabs.github.io/Fast-dLLM
Fast-dLLM: Training-free Acceleration of Diffusion Language Models
By
–
