1 paper
Jianing Yang, Yusuke Fujita, Yui Sudo
Spoken dialog systems with cascaded ASR-LLM-TTS modules retain strong LLM intelligence, but VAD segmentation often forces half-duplex turns and brittle control. On the other hand,…