13 citations · 17 across the 4 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2024
GLOBE: A High-quality English Corpus with Global Accents for Zero-shot Speaker Adaptive Text-to-Speech
Wenbin Wang, Yang Song, Sanjay Jha
This paper introduces GLOBE, a high-quality English corpus with worldwide accents, specifically designed to address the limitations of current zero-shot speaker adaptive Text-to-Sp…
cs.SD2024★ 13 cited
USAT: A Universal Speaker-Adaptive Text-to-Speech Approach
Wenbin Wang, Yang Song, Sanjay Jha
Conventional text-to-speech (TTS) research has predominantly focused on enhancing the quality of synthesized speech for speakers in the training dataset. The challenge of synthesiz…
cs.SD2023
Generalizable Zero-Shot Speaker Adaptive Speech Synthesis with Disentangled Representations
Wenbin Wang, Yang Song, Sanjay Jha
While most research into speech synthesis has focused on synthesizing high-quality speech for in-dataset speakers, an equally essential yet unsolved problem is synthesizing speech…