51 citations · 104 across the 20 of their papers we have counts for
3 papers · 1 filter
Towards Speaker Age Estimation with Label Distribution Learning
Shijing Si, Jianzong Wang, Junqing Peng +1
Existing methods for speaker age estimation usually treat it as a multi-class classification or a regression problem. However, precise age identification remains a challenge due to…
Speech2Video: Cross-Modal Distillation for Speech to Video Generation
Shijing Si, Jianzong Wang, Xiaoyang Qu +4
This paper investigates a novel task of talking face video generation solely from speeches. The speech-to-video generation technique can spark interesting applications in entertain…
Variational Information Bottleneck for Effective Low-resource Audio Classification
Shijing Si, Jianzong Wang, Huiming Sun +6
Large-scale deep neural networks (DNNs) such as convolutional neural networks (CNNs) have achieved impressive performance in audio classification for their powerful capacity and st…