Story

arxiv_cs_lg ยท Aug 13, 2026 ยท paper

Source brief

DARTree: Speculative Diffusion Decoding with Autoregressive Draft Trees

arxiv.orgAug 13, 2026
original source linked

In brief

Speculative decoding losslessly accelerates autoregressive language models by verifying multiple draft tokens in parallel. Diffusion-based drafters further reduce proposal latency by predicting an entire token block i...

Continue reading

Read the original at arxiv.org โ†’Open in live feedRead that dayโ€™s brief

Earlier in this thread 4 items