58 citations · 130 across the 67 of their papers we have counts for
64 papers
SVLDL: Improved Speaker Age Estimation Using Selective Variance Label Distribution Learning
Zuheng Kang, Jianzong Wang, Junqing Peng +1
Estimating age from a single speech is a classic and challenging topic. Although Label Distribution Learning (LDL) can represent adjacent indistinguishable ages well, the uncertain…
Learning Invariant Representation and Risk Minimized for Unsupervised Accent Domain Adaptation
Chendong Zhao, Jianzong Wang, Xiaoyang Qu +2
Unsupervised representation learning for speech audios attained impressive performances for speech recognition tasks, particularly when annotated speech is limited. However, the un…
Linguistic-Enhanced Transformer with CTC Embedding for Speech Recognition
Xulong Zhang, Jianzong Wang, Ning Cheng +3
The recent emergence of joint CTC-Attention model shows significant improvement in automatic speech recognition (ASR). The improvement largely lies in the modeling of linguistic in…
Improving Imbalanced Text Classification with Dynamic Curriculum Learning
Xulong Zhang, Jianzong Wang, Ning Cheng +1
Recent advances in pre-trained language models have improved the performance for text classification tasks. However, little attention is paid to the priority scheduling strategy on…
Semi-Supervised Learning Based on Reference Model for Low-resource TTS
Xulong Zhang, Jianzong Wang, Ning Cheng +1
Most previous neural text-to-speech (TTS) methods are mainly based on supervised learning methods, which means they depend on a large training dataset and hard to achieve comparabl…
MetaSpeech: Speech Effects Switch Along with Environment for Metaverse
Xulong Zhang, Jianzong Wang, Ning Cheng +1
Metaverse expands the physical world to a new dimension, and the physical environment and Metaverse environment can be directly connected and entered. Voice is an indispensable com…