158 citations · 419 across the 48 of their papers we have counts for
4 papers · 2 filters
Detecting Adversarial Attacks On Audiovisual Speech Recognition
Pingchuan Ma, Stavros Petridis, Maja Pantic
Adversarial attacks pose a threat to deep learning models. However, research on adversarial detection methods, especially in the multi-modal domain, is very limited. In this work,…
Towards Pose-invariant Lip-Reading
Shiyang Cheng, Pingchuan Ma, Georgios Tzimiropoulos +4
Lip-reading models have been significantly improved recently thanks to powerful deep learning architectures. However, most works focused on frontal or near frontal views of the mou…
Realistic Speech-Driven Facial Animation with GANs
Konstantinos Vougioukas, Stavros Petridis, Maja Pantic
Speech-driven facial animation is the process that automatically synthesizes talking characters based on speech signals. The majority of work in this domain creates a mapping from…
End-to-End Visual Speech Recognition for Small-Scale Datasets
Stavros Petridis, Yujiang Wang, Pingchuan Ma +2
Visual speech recognition models traditionally consist of two stages, feature extraction and classification. Several deep learning approaches have been recently presented aiming to…