158 citations · 419 across the 48 of their papers we have counts for
Showing 2022 · cs.SDShow all
2 papers · 2 filters
cs.SD2022★ 1 cited
LA-VocE: Low-SNR Audio-visual Speech Enhancement using Neural Vocoders
Rodrigo Mira, Buye Xu, Jacob Donley +4
Audio-visual speech enhancement aims to extract clean speech from a noisy environment by leveraging not only the audio itself but also the target speaker's lip movements. This appr…
cs.SD2022★ 1 cited
SVTS: Scalable Video-to-Speech Synthesis
Rodrigo Mira, Alexandros Haliassos, Stavros Petridis +2
Video-to-speech synthesis (also known as lip-to-speech) refers to the translation of silent lip movements into the corresponding audio. This task has received an increasing amount…