3 citations · 3 across the 1 of their papers we have counts for
3 papers
cs.SD2022
DGC-vector: A new speaker embedding for zero-shot voice conversion
Ruitong Xiao, Haitong Zhang, Yue Lin
Recently, more and more zero-shot voice conversion algorithms have been proposed. As a fundamental part of zero-shot voice conversion, speaker embeddings are the key to improving t…
cs.SD2022
Improve few-shot voice cloning using multi-modal learning
Haitong Zhang, Yue Lin
Recently, few-shot voice cloning has achieved a significant improvement. However, most models for few-shot voice cloning are single-modal, and multi-modal few-shot voice cloning ha…
cs.CL2021★ 3 cited
Revisiting IPA-based Cross-lingual Text-to-speech
Haitong Zhang, Haoyue Zhan, Yang Zhang +2
International Phonetic Alphabet (IPA) has been widely used in cross-lingual text-to-speech (TTS) to achieve cross-lingual voice cloning (CL VC). However, IPA itself has been unders…