Showing cs.LGShow all
3 papers · 1 filter
cs.LG2024
Focused Discriminative Training For Streaming CTC-Trained Automatic Speech Recognition Models
Adnan Haider, Xingyu Na, Erik McDermott +3
This paper introduces a novel training framework called Focused Discriminative Training (FDT) to further improve streaming word-piece end-to-end (E2E) automatic speech recognition…
cs.LG2023
Conformer-Based Speech Recognition On Extreme Edge-Computing Devices
Mingbin Xu, Alex Jin, Sicheng Wang +8
With increasingly more powerful compute capabilities and resources in today's devices, traditionally compute-intensive automatic speech recognition (ASR) has been moving from the c…
cs.LG2023
Personalization of CTC-based End-to-End Speech Recognition Using Pronunciation-Driven Subword Tokenization
Zhihong Lei, Ernest Pusateri, Shiyi Han +8
Recent advances in deep learning and automatic speech recognition have improved the accuracy of end-to-end speech recognition systems, but recognition of personal content such as c…