5 citations · 10 across the 3 of their papers we have counts for
3 papers
eess.AS2023★ 1 cited
Adapting Large Language Model with Speech for Fully Formatted End-to-End Speech Recognition
Shaoshi Ling, Yuxuan Hu, Shuangbei Qian +5
Most end-to-end (E2E) speech recognition models are composed of encoder and decoder blocks that perform acoustic and language modeling functions. Pretrained large language models (…
cs.CL2022★ 5 cited
Mask the Correct Tokens: An Embarrassingly Simple Approach for Error Correction
Kai Shen, Yichong Leng, Xu Tan +4
Text error correction aims to correct the errors in text sequences such as those typed by humans or generated by speech recognition models. Previous error correction methods usuall…
eess.AS2020★ 4 cited
An End-to-end Architecture of Online Multi-channel Speech Separation
Jian Wu, Zhuo Chen, Jinyu Li +5
Multi-speaker speech recognition has been one of the keychallenges in conversation transcription as it breaks the singleactive speaker assumption employed by most state-of-the-arts…