1 paper · 1 filter
Muyang Du, Jason Roche, Junjie Lai
Streaming text-to-speech synthesis in cascaded LLM-TTS systems still faces latency challenges as most TTS models require full context before initiating generation. We present S5-TT…