1 citations · 1 across the 3 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2023
Towards Streaming Speech-to-Avatar Synthesis
Tejas S. Prabhune, Peter Wu, Bohan Yu +1
Streaming speech-to-avatar synthesis creates real-time animations for a virtual character from audio data. Accurate avatar representations of speech are important for the visualiza…
cs.SD2023
Towards an Interpretable Representation of Speaker Identity via Perceptual Voice Qualities
Robin Netzorg, Bohan Yu, Andrea Guzman +3
Unlike other data modalities such as text and vision, speech does not lend itself to easy interpretation. While lay people can understand how to describe an image or sentence via p…