49 citations · 158 across the 25 of their papers we have counts for
25 papers
Triplet Knowledge Distillation
Xijun Wang, Dongyang Liu, Meina Kan +3
In Knowledge Distillation, the teacher is generally much larger than the student, making the solution of the teacher likely to be difficult for the student to learn. To ease the mi…
FusionFormer: Fusing Operations in Transformer for Efficient Streaming Speech Recognition
Xingchen Song, Di Wu, Binbin Zhang +8
The recently proposed Conformer architecture which combines convolution with attention to capture both local and global dependencies has become the \textit{de facto} backbone model…
A Character-level Span-based Model for Mandarin Prosodic Structure Prediction
Xueyuan Chen, Changhe Song, Yixuan Zhou +4
The accuracy of prosodic structure prediction is crucial to the naturalness of synthesized speech in Mandarin text-to-speech system, but now is limited by widely-used sequence-to-s…
Syntax-Aware Network for Handwritten Mathematical Expression Recognition
Ye Yuan, Xiao Liu, Wondimu Dikubab +4
Handwritten mathematical expression recognition (HMER) is a challenging task that has many potential applications. Recent methods for HMER have achieved outstanding performance wit…
BERT-LID: Leveraging BERT to Improve Spoken Language Identification
Yuting Nie, Junhong Zhao, Wei-Qiang Zhang +1
Language identification is the task of automatically determining the identity of a language conveyed by a spoken segment. It has a profound impact on the multilingual interoperabil…
Orthogonal Jacobian Regularization for Unsupervised Disentanglement in Image Generation
Yuxiang Wei, Yupeng Shi, Xiao Liu +4
Unsupervised disentanglement learning is a crucial issue for understanding and exploiting deep generative models. Recently, SeFa tries to find latent disentangled directions by per…