42 citations · 66 across the 16 of their papers we have counts for
Showing 2026Show all
3 papers · 1 filter
cs.CL2026
Low Perplexity is Repetition: A One-Dimensional Self-Conditioning Attractor in Continuous Diffusion LMs
Shuai Zhang, Zijie Chen, Hongliang He +2
Continuous diffusion language models such as ELF report record-low generative perplexity (Gen-PPL). We find a catch: these models repeat far more than human text, and Gen-PPL rewar…
cs.CL2026
Cost-Aware Diffusion Draft Trees for Speculative Decoding
Shuai Zhang, Huachuan Qiu, Hongliang He +1
Speculative decoding accelerates inference by having a lightweight drafter propose tokens verified in parallel by the target language model. Block diffusion drafters such as DFlash…
cs.LG2026
LLaDA2.1: Speeding Up Text Diffusion via Token Editing
Tiwei Bie, Maosong Cao, Xiang Cao +47
While LLaDA2.0 showcased the scaling potential of 100B-level block-diffusion models and their inherent parallelization, the delicate equilibrium between decoding speed and generati…