Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Deep CLAS: Deep Contextual Listen, Attend and Spell
Mengzhi Wang, Shifu Xiong, Genshun Wan +3
Contextual-LAS (CLAS) has been shown effective in improving Automatic Speech Recognition (ASR) of rare words. It relies on phrase-level contextual modeling and attention-based rele…
cs.CL2024
Lightweight Transducer Based on Frame-Level Criterion
Genshun Wan, Mengzhi Wang, Tingzhi Mao +2
The transducer model trained based on sequence-level criterion requires a lot of memory due to the generation of the large probability matrix. We proposed a lightweight transducer…