59 citations · 87 across the 10 of their papers we have counts for
27 papers
SS-VAERR: Self-Supervised Apparent Emotional Reaction Recognition from Video
Marija Jegorova, Stavros Petridis, Maja Pantic
This work focuses on the apparent emotional reaction recognition (AERR) from the video-only input, conducted in a self-supervised fashion. The network is first pre-trained on diffe…
Training Strategies for Improved Lip-reading
Pingchuan Ma, Yujiang Wang, Stavros Petridis +2
Several training strategies and temporal models have been recently proposed for isolated word lip-reading in a series of independent works. However, the potential of combining the…
Domain Generalisation for Apparent Emotional Facial Expression Recognition across Age-Groups
Rafael Poyiadzi, Jie Shen, Stavros Petridis +2
Apparent emotional facial expression recognition has attracted a lot of research attention recently. However, the majority of approaches ignore age differences and train a generic…
LiRA: Learning Visual Speech Representations from Audio through Self-supervision
Pingchuan Ma, Rodrigo Mira, Stavros Petridis +2
The large amount of audiovisual content being shared online today has drawn substantial attention to the prospect of audiovisual self-supervised learning. Recent works have focused…
DINO: A Conditional Energy-Based GAN for Domain Translation
Konstantinos Vougioukas, Stavros Petridis, Maja Pantic
Domain translation is the process of transforming data from one domain to another while preserving the common semantics. Some of the most popular domain translation systems are bas…
End-to-end Audio-visual Speech Recognition with Conformers
Pingchuan Ma, Stavros Petridis, Maja Pantic
In this work, we present a hybrid CTC/Attention model based on a ResNet-18 and Convolution-augmented transformer (Conformer), that can be trained in an end-to-end manner. In partic…