2 papers
cs.SD2023
Towards Streaming Speech-to-Avatar Synthesis
Tejas S. Prabhune, Peter Wu, Bohan Yu +1
Streaming speech-to-avatar synthesis creates real-time animations for a virtual character from audio data. Accurate avatar representations of speech are important for the visualiza…
cs.SD2023
CiwaGAN: Articulatory information exchange
Gašper Beguš, Thomas Lu, Alan Zhou +2
Humans encode information into sounds by controlling articulators and decode information from sounds using the auditory apparatus. This paper introduces CiwaGAN, a model of human s…