Showing eess.ASShow all
2 papers · 1 filter
eess.AS2024
PRESENT: Zero-Shot Text-to-Prosody Control
Perry Lam, Huayun Zhang, Nancy F. Chen +2
Current strategies for achieving fine-grained prosody control in speech synthesis entail extracting additional style embeddings or adopting more complex architectures. To enable ze…
eess.AS2024
SNIPER Training: Single-Shot Sparse Training for Text-to-Speech
Perry Lam, Huayun Zhang, Nancy F. Chen +2
Text-to-speech (TTS) models have achieved remarkable naturalness in recent years, yet like most deep neural models, they have more parameters than necessary. Sparse TTS models can…