Showing eess.ASShow all
3 papers · 1 filter
eess.AS2024
Optimizing Byte-level Representation for End-to-end ASR
Roger Hsiao, Liuhui Deng, Erik McDermott +2
We propose a novel approach to optimizing a byte-level representation for end-to-end automatic speech recognition (ASR). Byte-level representation is often used by large scale mult…
eess.AS2020
Online Automatic Speech Recognition with Listen, Attend and Spell Model
Roger Hsiao, Dogan Can, Tim Ng +2
The Listen, Attend and Spell (LAS) model and other attention-based automatic speech recognition (ASR) models have known limitations when operated in a fully online mode. In this pa…
eess.AS2019
Towards Adapting NMF Dictionaries Using Total Variability Modeling for Noise-Robust Acoustic Features
Kunal Dhawan, Colin Vaz, Ruchir Travadi +1
We propose an algorithm to extract noise-robust acoustic features from noisy speech. We use Total Variability Modeling in combination with Non-negative Matrix Factorization (NMF) t…