2 papers
cs.CL2026
Dynin-Omni: Omnimodal Unified Large Diffusion Language Model
Jaeik Kim, Woojin Kim, Jihwan Hong +8
We present Dynin-Omni, the first masked-diffusion-based omnimodal foundation model that unifies text, image, and speech understanding and generation, together with video understand…
eess.AS2025
Hybrid Decoding: Rapid Pass and Selective Detailed Correction for Sequence Models
Yunkyu Lim, Jihwan Park, Hyung Yong Kim +2
Recently, Transformer-based encoder-decoder models have demonstrated strong performance in multilingual speech recognition. However, the decoder's autoregressive nature and large s…