2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.SD2024
Streaming Decoder-Only Automatic Speech Recognition with Discrete Speech Units: A Pilot Study
Peikun Chen, Sining Sun, Changhao Shan +2
Unified speech-text models like SpeechGPT, VioLA, and AudioPaLM have shown impressive performance across various speech-related tasks, especially in Automatic Speech Recognition (A…
cs.CL2024
Skipformer: A Skip-and-Recover Strategy for Efficient Speech Recognition
Wenjing Zhu, Sining Sun, Changhao Shan +2
Conformer-based attention models have become the de facto backbone model for Automatic Speech Recognition tasks. A blank symbol is usually introduced to align the input and output…
cs.SD2023★ 2 cited
Key Frame Mechanism For Efficient Conformer Based End-to-end Speech Recognition
Peng Fan, Changhao Shan, Sining Sun +2
Recently, Conformer as a backbone network for end-to-end automatic speech recognition achieved state-of-the-art performance. The Conformer block leverages a self-attention mechanis…