3 citations · 6 across the 2 of their papers we have counts for
4 papers
Distilling Knowledge from Pre-trained Language Models via Text Smoothing
Xing Wu, Yibing Liu, Xiangyang Zhou +1
This paper studies compressing pre-trained language models, like BERT (Devlin et al.,2019), via teacher-student knowledge distillation. Previous works usually force the student mod…
Data Augmentation for Copy-Mechanism in Dialogue State Tracking
Xiaohui Song, Liangjun Zang, Yipeng Su +3
While several state-of-the-art approaches to dialogue state tracking (DST) have shown promising performances on several benchmarks, there is still a significant performance gap bet…
TransSent: Towards Generation of Structured Sentences with Discourse Marker
Xing Wu, Dongjun Wei, Liangjun Zang +2
Structured sentences are important expressions in human writings and dialogues. Previous works on neural text generation fused semantic and structural information by encoding the e…
"Mask and Infill" : Applying Masked Language Model to Sentiment Transfer
Xing Wu, Tao Zhang, Liangjun Zang +2
This paper focuses on the task of sentiment transfer on non-parallel text, which modifies sentiment attributes (e.g., positive or negative) of sentences while preserving their attr…