1 paper
Tian-Hao Zhang, Dinghao Zhou, Guiping Zhong +2
RNN-T models are widely used in ASR, which rely on the RNN-T loss to achieve length alignment between input audio and target sequence. However, the implementation complexity and th…