11 citations · 20 across the 5 of their papers we have counts for
5 papers
Self-critical Sequence Training for Automatic Speech Recognition
Chen Chen, Yuchen Hu, Nana Hou +3
Although automatic speech recognition (ASR) task has gained remarkable success by sequence-to-sequence models, there are two main mismatches between its training and testing that m…
Interactive Audio-text Representation for Automated Audio Captioning with Contrastive Learning
Chen Chen, Nana Hou, Yuchen Hu +3
Automated Audio captioning (AAC) is a cross-modal task that generates natural language to describe the content of input audio. Most prior works usually extract single-modality acou…
Noise-robust Speech Recognition with 10 Minutes Unparalleled In-domain Data
Chen Chen, Nana Hou, Yuchen Hu +2
Noise-robust speech recognition systems require large amounts of training data including noisy speech data and corresponding transcripts to achieve state-of-the-art performances in…
Progressive Continual Learning for Spoken Keyword Spotting
Yizheng Huang, Nana Hou, Nancy F. Chen
Catastrophic forgetting is a thorny challenge when updating keyword spotting (KWS) models after deployment. To tackle such challenges, we propose a progressive continual learning s…
Multitask-Based Joint Learning Approach To Robust ASR For Radio Communication Speech
Duo Ma, Nana Hou, Van Tung Pham +2
To realize robust end-to-end Automatic Speech Recognition(E2E ASR) under radio communication condition, we propose a multitask-based method to joint train a Speech Enhancement (SE)…