5 citations · 6 across the 8 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2025
A Non-autoregressive Model for Joint STT and TTS
Vishal Sunder, Brian Kingsbury, George Saon +5
In this paper, we take a step towards jointly modeling automatic speech recognition (STT) and speech synthesis (TTS) in a fully non-autoregressive way. We develop a novel multimoda…
cs.SD2022★ 5 cited
Speech Emotion Recognition using Self-Supervised Features
Edmilson Morais, Ron Hoory, Weizhong Zhu +3
Self-supervised pre-trained features have consistently delivered state-of-art results in the field of natural language processing (NLP); however, their merits in the field of speec…