3 papers
cs.SD2023
Audio-visual video-to-speech synthesis with synthesized input audio
Triantafyllos Kefalas, Yannis Panagakis, Maja Pantic
Video-to-speech synthesis involves reconstructing the speech signal of a speaker from a silent video. The implicit assumption of this task is that the sound signal is either missin…
cs.SD2023
Large-scale unsupervised audio pre-training for video-to-speech synthesis
Triantafyllos Kefalas, Yannis Panagakis, Maja Pantic
Video-to-speech synthesis is the task of reconstructing the speech signal from a silent video of a speaker. Most established approaches to date involve a two-step process, whereby…
cs.LG2019
Speech-driven facial animation using polynomial fusion of features
Triantafyllos Kefalas, Konstantinos Vougioukas, Yannis Panagakis +3
Speech-driven facial animation involves using a speech signal to generate realistic videos of talking faces. Recent deep learning approaches to facial synthesis rely on extracting…