Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Speculative Correction: Draft-then-Refine Decoding for Diffusion Language Models
Brian K Chen, Chong Wu, Kenji Kawaguchi
Diffusion language models (DLMs) can revise tokens bidirectionally, but standard decoding procedures often adapt them to left-to-right generation by producing text block by block.…
cs.CL2026
PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention
Haonan Wang, Brian Chen, Siquan Li +4
Parameter-Efficient Fine-Tuning (PEFT) methods have become crucial for rapidly adapting large language models (LLMs) to downstream tasks. Prefix-Tuning, an early and effective PEFT…