5 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.CL2020★ 3 cited
Distilling Knowledge from Pre-trained Language Models via Text Smoothing
Xing Wu, Yibing Liu, Xiangyang Zhou +1
This paper studies compressing pre-trained language models, like BERT (Devlin et al.,2019), via teacher-student knowledge distillation. Previous works usually force the student mod…
cs.CL2019★ 5 cited
How to Evaluate the Next System: Automatic Dialogue Evaluation from the Perspective of Continual Learning
Lu Li, Zhongheng He, Xiangyang Zhou +1
Automatic dialogue evaluation plays a crucial role in open-domain dialogue research. Previous works train neural networks with limited annotation for conducting automatic dialogue…