1 paper · 1 filter
YoungChae Kim, Da-Hee Yang, Joon-Hyuk Chang
Large language model (LLM)-based audio-visual speech recognition (AVSR) systems are robust under noise. Contrastive decoding (CD), originally introduced to stabilize LLM generation…