9 citations · 11 across the 3 of their papers we have counts for
4 papers
Cloning one's voice using very limited data in the wild
Dongyang Dai, Yuanzhe Chen, Li Chen +6
With the increasing popularity of speech synthesis products, the industry has put forward more requirements for personalized speech synthesis: (1) How to use low-resource, easily a…
The ByteDance Speaker Diarization System for the VoxCeleb Speaker Recognition Challenge 2021
Keke Wang, Xudong Mao, Hao Wu +4
This paper describes the ByteDance speaker diarization system for the fourth track of the VoxCeleb Speaker Recognition Challenge 2021 (VoxSRC-21). The VoxSRC-21 provides both the d…
Speech enhancement with weakly labelled data from AudioSet
Qiuqiang Kong, Haohe Liu, Xingjian Du +3
Speech enhancement is a task to improve the intelligibility and perceptual quality of degraded speech signal. Recently, neural networks based methods have been applied to speech en…
Noise Robust TTS for Low Resource Speakers using Pre-trained Model and Speech Enhancement
Dongyang Dai, Li Chen, Yuping Wang +5
With the popularity of deep neural network, speech synthesis task has achieved significant improvements based on the end-to-end encoder-decoder framework in the recent days. More a…