2 citations · 2 across the 4 of their papers we have counts for
1 paper · 1 filter
Georu Lee, Seungwon Jeong, Hoki Kim +2
Recent masked diffusion language models (MDLMs), such as LLaDA and Dream, have achieved performance comparable to autoregressive large language models. Unlike autoregressive models…