2 papers
cs.SD2025
A Non-autoregressive Model for Joint STT and TTS
Vishal Sunder, Brian Kingsbury, George Saon +5
In this paper, we take a step towards jointly modeling automatic speech recognition (STT) and speech synthesis (TTS) in a fully non-autoregressive way. We develop a novel multimoda…
cs.LG2023
Soft Random Sampling: A Theoretical and Empirical Analysis
Xiaodong Cui, Ashish Mittal, Songtao Lu +3
Soft random sampling (SRS) is a simple yet effective approach for efficient training of large-scale deep neural networks when dealing with massive data. SRS selects a subset unifor…