4 citations · 4 across the 1 of their papers we have counts for
2 papers
eess.AS2020★ 4 cited
Research on Modeling Units of Transformer Transducer for Mandarin Speech Recognition
Li Fu, Xiaoxiao Li, Libo Zi
Modeling unit and model architecture are two key factors of Recurrent Neural Network Transducer (RNN-T) in end-to-end speech recognition. To improve the performance of RNN-T for Ma…
cs.CV2019
Learning Enhanced Resolution-wise features for Human Pose Estimation
Kun Zhang, Peng He, Ping Yao +6
Recently, multi-resolution networks (such as Hourglass, CPN, HRNet, etc.) have achieved significant performance on pose estimation by combining feature maps of various resolutions.…