6 citations · 6 across the 1 of their papers we have counts for
4 papers
SpeechNet: A Universal Modularized Model for Speech Processing Tasks
Yi-Chen Chen, Po-Han Chi, Shu-wen Yang +7
There is a wide variety of speech processing tasks ranging from extracting content information from speech signals to generating speech signals. For different tasks, model networks…
Investigating on Incorporating Pretrained and Learnable Speaker Representations for Multi-Speaker Multi-Style Text-to-Speech
Chung-Ming Chien, Jheng-Hao Lin, Chien-yu Huang +2
The few-shot multi-speaker multi-style voice cloning task is to synthesize utterances with voice and speaking style similar to a reference speaker given only a few reference sample…
S2VC: A Framework for Any-to-Any Voice Conversion with Self-Supervised Pretrained Representations
Jheng-hao Lin, Yist Y. Lin, Chung-Ming Chien +1
Any-to-any voice conversion (VC) aims to convert the timbre of utterances from and to any speakers seen or unseen during training. Various any-to-any VC approaches have been propos…
How Far Are We from Robust Voice Conversion: A Survey
Tzu-hsien Huang, Jheng-hao Lin, Chien-yu Huang +1
Voice conversion technologies have been greatly improved in recent years with the help of deep learning, but their capabilities of producing natural sounding utterances in differen…