activity
20202025
most citedApplying Wav2vec2.0 to Speech Recognition in Various Low-resource Languages

58 citations · 130 across the 67 of their papers we have counts for

collaborators

64 papers

cs.SD2022

SVLDL: Improved Speaker Age Estimation Using Selective Variance Label Distribution Learning

Zuheng Kang, Jianzong Wang, Junqing Peng +1

Estimating age from a single speech is a classic and challenging topic. Although Label Distribution Learning (LDL) can represent adjacent indistinguishable ages well, the uncertain…

cs.SD2022

Learning Invariant Representation and Risk Minimized for Unsupervised Accent Domain Adaptation

Chendong Zhao, Jianzong Wang, Xiaoyang Qu +2

Unsupervised representation learning for speech audios attained impressive performances for speech recognition tasks, particularly when annotated speech is limited. However, the un…

cs.CL2022

Linguistic-Enhanced Transformer with CTC Embedding for Speech Recognition

Xulong Zhang, Jianzong Wang, Ning Cheng +3

The recent emergence of joint CTC-Attention model shows significant improvement in automatic speech recognition (ASR). The improvement largely lies in the modeling of linguistic in…

cs.CL2022

Improving Imbalanced Text Classification with Dynamic Curriculum Learning

Xulong Zhang, Jianzong Wang, Ning Cheng +1

Recent advances in pre-trained language models have improved the performance for text classification tasks. However, little attention is paid to the priority scheduling strategy on…

cs.SD2022

Semi-Supervised Learning Based on Reference Model for Low-resource TTS

Xulong Zhang, Jianzong Wang, Ning Cheng +1

Most previous neural text-to-speech (TTS) methods are mainly based on supervised learning methods, which means they depend on a large training dataset and hard to achieve comparabl…

cs.SD2022

MetaSpeech: Speech Effects Switch Along with Environment for Metaverse

Xulong Zhang, Jianzong Wang, Ning Cheng +1

Metaverse expands the physical world to a new dimension, and the physical environment and Metaverse environment can be directly connected and entered. Voice is an indispensable com…