3 papers
cs.LG2026
Simple Self-Conditioning Adaptation for Masked Diffusion Models
Michael Cardei, Huu Binh Ta, Ferdinando Fioretto
Masked diffusion models (MDMs) generate discrete sequences by iterative denoising under an absorbing masking process. In standard masked diffusion, if a token remains masked after…
cs.LG2026
Selective Sinkhorn Routing for Improved Sparse Mixture of Experts
Duc Anh Nguyen, Huu Binh Ta, Nhuan Le Duc +2
Sparse Mixture-of-Experts (SMoE) models are scalable and computationally efficient, enabling large increases in model capacity with limited inference overhead. Existing SMoE method…
cs.LG2026
Search-Augmented Masked Diffusion Models for Constrained Generation
Huu Binh Ta, Michael Cardei, Alvaro Velasquez +1
Discrete diffusion models generate sequences by iteratively denoising samples corrupted by categorical noise, offering an appealing alternative to autoregressive decoding for struc…