2 citations · 3 across the 2 of their papers we have counts for
3 papers
cs.SD2022★ 1 cited
NeuFA: Neural Network Based End-to-End Forced Alignment with Bidirectional Attention Mechanism
Jingbei Li, Yi Meng, Zhiyong Wu +4
Although deep learning and end-to-end models have been widely used and shown superiority in automatic speech recognition (ASR) and text-to-speech (TTS) synthesis, state-of-the-art…
eess.AS2022★ 2 cited
Building Synthetic Speaker Profiles in Text-to-Speech Systems
Jie Pu, Yixiong Meng, Oguz Elibol
The diversity of speaker profiles in multi-speaker TTS systems is a crucial aspect of its performance, as it measures how many different speaker profiles TTS systems could possibly…
cs.LG2021
SynthASR: Unlocking Synthetic Data for Speech Recognition
Amin Fazel, Wei Yang, Yulan Liu +4
End-to-end (E2E) automatic speech recognition (ASR) models have recently demonstrated superior performance over the traditional hybrid ASR models. Training an E2E ASR model require…