Story
arxiv_cs_ai ยท Apr 29, 2026 ยท paper
arxiv.orgApr 29, 2026
original source linked
In brief
Diffusion large language models (dLLMs) offer parallel decoding and bidirectional context, but state-of-the-art dLLMs require billions of parameters for competitive performance. While existing distillation methods for...
Continue reading