1 paper
Shengfan Shen, Di Wu, Xingchen Song +5
Expressive speech synthesis for voice assistants requires flexible style control that adapts to explicit requests and broader interaction context. We propose Harness TTS, a lightwe…