From the 1 of 1 linked paper with an AI index.
1 paper
Daehoon Gwak, Minhyung Lee, Junwoo Park +1
The paper surveys methods for speeding up inference of masked diffusion large language models by categorizing algorithmic, architectural, and system-level acceleration techniques a…