Authors

Xing Niu

Publications 1

Preprint Open access

DEdit: Iterative Draft Editing for Speculative Decoding

Longxuan Yu, Bingsen Chen, Peng Shi et al. · 2026

Speculative decoding accelerates autoregressive LLMs by having a lightweight drafter propose tokens that the target model verifies in parallel. Diffusion-based drafters further reduce drafting latency by proposing multiple tokens at once. However, these tokens are predicted independently, so a single early error causes …

Co-authors