1 citations · 1 across the 2 of their papers we have counts for
2 papers
eess.AS2025
An Exploration of ECAPA-TDNN and x-vector Speaker Representations in Zero-shot Multi-speaker TTS
Marie Kunešová, Zdeněk Hanzlíček, Jindřich Matoušek
Zero-shot multi-speaker text-to-speech (TTS) systems rely on speaker embeddings to synthesize speech in the voice of an unseen speaker, using only a short reference utterance. Whil…
cs.SD2024★ 1 cited
Zero-Shot vs. Few-Shot Multi-Speaker TTS Using Pre-trained Czech SpeechT5 Model
Jan Lehečka, Zdeněk Hanzlíček, Jindřich Matoušek +1
In this paper, we experimented with the SpeechT5 model pre-trained on large-scale datasets. We pre-trained the foundation model from scratch and fine-tuned it on a large-scale robu…