2 papers
cs.SD2025
NaturalL2S: End-to-End High-quality Multispeaker Lip-to-Speech Synthesis with Differential Digital Signal Processing
Yifan Liang, Fangkun Liu, Andong Li +2
Recent advancements in visual speech recognition (VSR) have promoted progress in lip-to-speech synthesis, where pre-trained VSR models enhance the intelligibility of synthesized sp…
cs.SD2025
Neural Vocoders as Speech Enhancers
Andong Li, Zhihang Sun, Fengyuan Hao +2
Speech enhancement (SE) and neural vocoding are traditionally viewed as separate tasks. In this work, we observe them under a common thread: the rank behavior of these processes. T…