1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2023
Emotional Talking Head Generation based on Memory-Sharing and Attention-Augmented Networks
Jianrong Wang, Yaxin Zhao, Li Liu +3
Given an audio clip and a reference face image, the goal of the talking head generation is to generate a high-fidelity talking head video. Although some audio-driven methods of gen…
cs.SD2023
MAVD: The First Open Large-Scale Mandarin Audio-Visual Dataset with Depth Information
Jianrong Wang, Yuchen Huo, Li Liu +3
Audio-visual speech recognition (AVSR) gains increasing attention from researchers as an important part of human-computer interaction. However, the existing available Mandarin audi…
cs.MM2023★ 1 cited
Memory-augmented Contrastive Learning for Talking Head Generation
Jianrong Wang, Yaxin Zhao, Li Liu +4
Given one reference facial image and a piece of speech as input, talking head generation aims to synthesize a realistic-looking talking head video. However, generating a lip-synchr…