11 citations · 28 across the 11 of their papers we have counts for
15 papers
Minimum word error training for non-autoregressive Transformer-based code-switching ASR
Yizhou Peng, Jicheng Zhang, Haihua Xu +2
Non-autoregressive end-to-end ASR framework might be potentially appropriate for code-switching recognition task thanks to its inherent property that present output token being ind…
Multitask-Based Joint Learning Approach To Robust ASR For Radio Communication Speech
Duo Ma, Nana Hou, Van Tung Pham +2
To realize robust end-to-end Automatic Speech Recognition(E2E ASR) under radio communication condition, we propose a multitask-based method to joint train a Speech Enhancement (SE)…
E2E-based Multi-task Learning Approach to Joint Speech and Accent Recognition
Jicheng Zhang, Yizhou Peng, Pham Van Tung +3
In this paper, we propose a single multi-task learning framework to perform End-to-End (E2E) speech recognition (ASR) and accent recognition (AR) simultaneously. The proposed frame…
Enriching Under-Represented Named-Entities To Improve Speech Recognition Performance
Tingzhi Mao, Yerbolat Khassanov, Van Tung Pham +4
Automatic speech recognition (ASR) for under-represented named-entity (UR-NE) is challenging due to such named-entities (NE) have insufficient instances and poor contextual coverag…
The NTU-AISG Text-to-speech System for Blizzard Challenge 2020
Haobo Zhang, Tingzhi Mao, Haihua Xu +1
We report our NTU-AISG Text-to-speech (TTS) entry systems for the Blizzard Challenge 2020 in this paper. There are two TTS tasks in this year's challenge, one is a Mandarin TTS tas…
Monolingual Data Selection Analysis for English-Mandarin Hybrid Code-switching Speech Recognition
Haobo Zhang, Haihua Xu, Van Tung Pham +2
In this paper, we conduct data selection analysis in building an English-Mandarin code-switching (CS) speech recognition (CSSR) system, which is aimed for a real CSSR contest in Ch…