139 citations · 207 across the 14 of their papers we have counts for
8 papers · 1 filter
Cross-domain Single-channel Speech Enhancement Model with Bi-projection Fusion Module for Noise-robust ASR
Fu-An Chao, Jeih-weih Hung, Berlin Chen
In recent decades, many studies have suggested that phase information is crucial for speech enhancement (SE), and time-domain single-channel speech enhancement techniques have show…
TENET: A Time-reversal Enhancement Network for Noise-robust ASR
Fu-An Chao, Shao-Wei Fan Jiang, Bi-Cheng Yan +2
Due to the unprecedented breakthroughs brought about by deep learning, speech enhancement (SE) techniques have been developed rapidly and play an important role prior to acoustic m…
The NTNU Taiwanese ASR System for Formosa Speech Recognition Challenge 2020
Fu-An Chao, Tien-Hong Lo, Shi-Yan Weng +3
This paper describes the NTNU ASR system participating in the Formosa Speech Recognition Challenge 2020 (FSR-2020) supported by the Formosa Speech in the Wild project (FSW). FSR-20…
End-to-End Mispronunciation Detection and Diagnosis From Raw Waveforms
Bi-Cheng Yan, Berlin Chen
Mispronunciation detection and diagnosis (MDD) is designed to identify pronunciation errors and provide instructive feedback to guide non-native language learners, which is a core…
Effective Decoder Masking for Transformer Based End-to-End Speech Recognition
Shi-Yan Weng, Berlin Chen
The attention-based encoder-decoder modeling paradigm has achieved promising results on a variety of speech processing tasks like automatic speech recognition (ASR), text-to-speech…
The NTNU System at the Interspeech 2020 Non-Native Children's Speech ASR Challenge
Tien-Hong Lo, Fu-An Chao, Shi-Yan Weng +1
This paper describes the NTNU ASR system participating in the Interspeech 2020 Non-Native Children's Speech ASR Challenge supported by the SIG-CHILD group of ISCA. This ASR shared…