5 citations · 10 across the 3 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2023★ 1 cited
Adapting Large Language Model with Speech for Fully Formatted End-to-End Speech Recognition
Shaoshi Ling, Yuxuan Hu, Shuangbei Qian +5
Most end-to-end (E2E) speech recognition models are composed of encoder and decoder blocks that perform acoustic and language modeling functions. Pretrained large language models (…
eess.AS2020★ 4 cited
An End-to-end Architecture of Online Multi-channel Speech Separation
Jian Wu, Zhuo Chen, Jinyu Li +5
Multi-speaker speech recognition has been one of the keychallenges in conversation transcription as it breaks the singleactive speaker assumption employed by most state-of-the-arts…