3 citations · 5 across the 3 of their papers we have counts for
3 papers
cs.CL2024
Bridging Speech and Text: Enhancing ASR with Pinyin-to-Character Pre-training in LLMs
Yang Yuhang, Peng Yizhou, Eng Siong Chng +1
The integration of large language models (LLMs) with pre-trained speech models has opened up new avenues in automatic speech recognition (ASR). While LLMs excel in multimodal under…
eess.AS2021★ 3 cited
Minimum word error training for non-autoregressive Transformer-based code-switching ASR
Yizhou Peng, Jicheng Zhang, Haihua Xu +2
Non-autoregressive end-to-end ASR framework might be potentially appropriate for code-switching recognition task thanks to its inherent property that present output token being ind…
eess.AS2021★ 2 cited
E2E-based Multi-task Learning Approach to Joint Speech and Accent Recognition
Jicheng Zhang, Yizhou Peng, Pham Van Tung +3
In this paper, we propose a single multi-task learning framework to perform End-to-End (E2E) speech recognition (ASR) and accent recognition (AR) simultaneously. The proposed frame…