59 citations · 104 across the 10 of their papers we have counts for
3 papers · 1 filter
Visually Guided Self Supervised Learning of Speech Representations
Abhinav Shukla, Konstantinos Vougioukas, Pingchuan Ma +2
Self supervised representation learning has recently attracted a lot of research interest for both the audio and visual modalities. However, most works typically focus on a particu…
Investigating the Lombard Effect Influence on End-to-End Audio-Visual Speech Recognition
Pingchuan Ma, Stavros Petridis, Maja Pantic
Several audio-visual speech recognition models have been recently proposed which aim to improve the robustness over audio-only models in the presence of noise. However, almost all…
Video-Driven Speech Reconstruction using Generative Adversarial Networks
Konstantinos Vougioukas, Pingchuan Ma, Stavros Petridis +1
Speech is a means of communication which relies on both audio and visual information. The absence of one modality can often lead to confusion or misinterpretation of information. I…