1 citations · 1 across the 4 of their papers we have counts for
4 papers
Emotional Talking Head Generation based on Memory-Sharing and Attention-Augmented Networks
Jianrong Wang, Yaxin Zhao, Li Liu +3
Given an audio clip and a reference face image, the goal of the talking head generation is to generate a high-fidelity talking head video. Although some audio-driven methods of gen…
MAVD: The First Open Large-Scale Mandarin Audio-Visual Dataset with Depth Information
Jianrong Wang, Yuchen Huo, Li Liu +3
Audio-visual speech recognition (AVSR) gains increasing attention from researchers as an important part of human-computer interaction. However, the existing available Mandarin audi…
Memory-augmented Contrastive Learning for Talking Head Generation
Jianrong Wang, Yaxin Zhao, Li Liu +4
Given one reference facial image and a piece of speech as input, talking head generation aims to synthesize a realistic-looking talking head video. However, generating a lip-synchr…
Two-Stream Joint-Training for Speaker Independent Acoustic-to-Articulatory Inversion
Jianrong Wang, Jinyu Liu, Li Liu +4
Acoustic-to-articulatory inversion (AAI) aims to estimate the parameters of articulators from speech audio. There are two common challenges in AAI, which are the limited data and t…