Masao Someki, Nicholas Eng, Yosuke Higuchi +1
Attention-based encoder-decoder models with autoregressive (AR) decoding have proven to be the dominant approach for automatic speech recognition (ASR) due to their superior accura…