2 papers
cs.CV2022
A Keypoint Based Enhancement Method for Audio Driven Free View Talking Head Synthesis
Yichen Han, Ya Li, Yingming Gao +3
Audio driven talking head synthesis is a challenging task that attracts increasing attention in recent years. Although existing methods based on 2D landmarks or 3D face models can…
cs.SD2022
ECAPA-TDNN for Multi-speaker Text-to-speech Synthesis
Jinlong Xue, Yayue Deng, Yichen Han +3
In recent years, neural network based methods for multi-speaker text-to-speech synthesis (TTS) have made significant progress. However, the current speaker encoder models used in t…