2 citations · 3 across the 2 of their papers we have counts for
2 papers
eess.AS2020★ 1 cited
PPSpeech: Phrase based Parallel End-to-End TTS System
Yahuan Cong, Ran Zhang, Jian Luan
Current end-to-end autoregressive TTS systems (e.g. Tacotron 2) have outperformed traditional parallel approaches on the quality of synthesized speech. However, they introduce new…
eess.AS2020★ 2 cited
Self-supervised learning for audio-visual speaker diarization
Yifan Ding, Yong Xu, Shi-Xiong Zhang +2
Speaker diarization, which is to find the speech segments of specific speakers, has been widely used in human-centered applications such as video conferences or human-computer inte…