2 papers
eess.AS2020
Online Automatic Speech Recognition with Listen, Attend and Spell Model
Roger Hsiao, Dogan Can, Tim Ng +2
The Listen, Attend and Spell (LAS) model and other attention-based automatic speech recognition (ASR) models have known limitations when operated in a fully online mode. In this pa…
cs.LG2019
SNDCNN: Self-normalizing deep CNNs with scaled exponential linear units for speech recognition
Zhen Huang, Tim Ng, Leo Liu +3
Very deep CNNs achieve state-of-the-art results in both computer vision and speech recognition, but are difficult to train. The most popular way to train very deep CNNs is to use s…