Publications (16)
Deep Semantic Manipulation of Facial Videos
Girish Kumar Solanki, Anastasios Roussos
Editing and manipulating facial features in videos is an interesting and important field of research with a plethora of applications, ranging from movie post-production and visual…
Real-Time Monocular 4D Face Reconstruction using the LSFM models
Mohammad Rami Koujan, Nikolai Dochev, Anastasios Roussos
4D face reconstruction from a single camera is a challenging task, especially when it is required to be performed in real time. We demonstrate a system of our own implementation th…
Visual Speech-Aware Perceptual 3D Facial Expression Reconstruction from Videos
Panagiotis P. Filntisis, George Retsinas, Foivos Paraperas-Papantoniou +3
The recent state of the art on monocular 3D face reconstruction from image data has made some impressive advancements, thanks to the advent of Deep Learning. However, it has mostly…
ReenactNet: Real-time Full Head Reenactment
Mohammad Rami Koujan, Michail Christos Doukas, Anastasios Roussos +1
Video-to-video synthesis is a challenging problem aiming at learning a translation function between a sequence of semantic maps and a photo-realistic video depicting the characteri…
Neural Text to Articulate Talk: Deep Text to Audiovisual Speech Synthesis achieving both Auditory and Photo-realism
Georgios Milis, Panagiotis P. Filntisis, Anastasios Roussos +1
Recent advances in deep learning for sequential data have given rise to fast and powerful models that produce realistic videos of talking humans. The state of the art in talking fa…
Reflections on Diversity: A Real-time Virtual Mirror for Inclusive 3D Face Transformations
Paraskevi Valergaki, Antonis Argyros, Giorgos Giannakakis +1
Real-time 3D face manipulation has significant applications in virtual reality, social media and human-computer interaction. This paper introduces a novel system, which we call Mir…
A Transformer-Based Framework for Greek Sign Language Production using Extended Skeletal Motion Representations
Chrysa Pratikaki, Panagiotis Filntisis, Athanasios Katsamanis +2
Sign Languages are the primary form of communication for Deaf communities across the world. To break the communication barriers between the Deaf and Hard-of-Hearing and the hearing…
3D Neural Sculpting (3DNS): Editing Neural Signed Distance Functions
Petros Tzathas, Petros Maragos, Anastasios Roussos
In recent years, implicit surface representations through neural networks that encode the signed distance have gained popularity and have achieved state-of-the-art results in vario…
Combining Facial Videos and Biosignals for Stress Estimation During Driving
Paraskevi Valergaki, Vassilis C. Nicodemou, Iason Oikonomidis +2
Reliable stress recognition is critical in applications such as medical monitoring and safety-critical systems, including real-world driving. While stress is commonly detected usin…
DeepFaceFlow: In-the-wild Dense 3D Facial Motion Estimation
Mohammad Rami Koujan, Anastasios Roussos, Stefanos Zafeiriou
Dense 3D facial motion capture from only monocular in-the-wild pairs of RGB images is a highly challenging problem with numerous applications, ranging from facial expression recogn…
SMIRK: 3D Facial Expressions through Analysis-by-Neural-Synthesis
George Retsinas, Panagiotis P. Filntisis, Radek Danecek +4
While existing methods for 3D face reconstruction from in-the-wild images excel at recovering the overall face shape, they commonly miss subtle, extreme, asymmetric, or rarely obse…
Real-time Facial Expression Recognition "In The Wild'' by Disentangling 3D Expression from Identity
Mohammad Rami Koujan, Luma Alharbawee, Giorgos Giannakakis +2
Human emotions analysis has been the focus of many studies, especially in the field of Affective Computing, and is important for many applications, e.g. human-computer intelligent…
Head2Head++: Deep Facial Attributes Re-Targeting
Michail Christos Doukas, Mohammad Rami Koujan, Viktoriia Sharmanska +1
Facial video re-targeting is a challenging problem aiming to modify the facial attributes of a target subject in a seamless manner by a driving monocular sequence. We leverage the…
Neural Emotion Director: Speech-preserving semantic control of facial expressions in "in-the-wild" videos
Foivos Paraperas Papantoniou, Panagiotis P. Filntisis, Petros Maragos +1
In this paper, we introduce a novel deep learning method for photo-realistic manipulation of the emotional state of actors in "in-the-wild" videos. The proposed method is based on…
Head2Head: Video-based Neural Head Synthesis
Mohammad Rami Koujan, Michail Christos Doukas, Anastasios Roussos +1
In this paper, we propose a novel machine learning architecture for facial reenactment. In particular, contrary to the model-based approaches or recent frame-based methods that use…
Neural Sign Reenactor: Deep Photorealistic Sign Language Retargeting
Christina O. Tze, Panagiotis P. Filntisis, Athanasia-Lida Dimou +2
In this paper, we introduce a neural rendering pipeline for transferring the facial expressions, head pose, and body movements of one person in a source video to another in a targe…