Showing eess.ASShow all
2 papers · 1 filter
eess.AS2023
Automatic Tuning of Loss Trade-offs without Hyper-parameter Search in End-to-End Zero-Shot Speech Synthesis
Seongyeon Park, Bohyung Kim, Tae-hyun Oh
Recently, zero-shot TTS and VC methods have gained attention due to their practicality of being able to generate voices even unseen during training. Among these methods, zero-shot…
eess.AS2023
Unsupervised Pre-Training For Data-Efficient Text-to-Speech On Low Resource Languages
Seongyeon Park, Myungseo Song, Bohyung Kim +1
Neural text-to-speech (TTS) models can synthesize natural human speech when trained on large amounts of transcribed speech. However, collecting such large-scale transcribed data is…