2 papers
cs.SD2026
TADA: A Generative Framework for Speech Modeling via Text-Acoustic Dual Alignment
Trung Dang, Sharath Rao, Ananya Gupta +6
Modern Text-to-Speech (TTS) systems increasingly leverage Large Language Model (LLM) architectures to achieve scalable, high-fidelity, zero-shot generation. However, these systems…
cs.SD2024
Zero-Shot Text-to-Speech from Continuous Text Streams
Trung Dang, David Aponte, Dung Tran +2
Existing zero-shot text-to-speech (TTS) systems are typically designed to process complete sentences and are constrained by the maximum duration for which they have been trained. H…