7 citations · 15 across the 5 of their papers we have counts for
5 papers
PreQuant: A Task-agnostic Quantization Approach for Pre-trained Language Models
Zhuocheng Gong, Jiahao Liu, Qifan Wang +6
While transformer-based pre-trained language models (PLMs) have dominated a number of NLP applications, these models are heavy to deploy and expensive to use. Therefore, effectivel…
MiniDisc: Minimal Distillation Schedule for Language Model Compression
Chen Zhang, Yang Yang, Qifan Wang +4
Recent studies have uncovered that language model distillation is less effective when facing a large capacity gap between the teacher and the student, and introduced teacher assist…
GNN-encoder: Learning a Dual-encoder Architecture via Graph Neural Networks for Dense Passage Retrieval
Jiduan Liu, Jiahao Liu, Yang Yang +4
Recently, retrieval models based on dense representations are dominant in passage retrieval tasks, due to their outstanding ability in terms of capturing semantics of input text co…
VIRT: Improving Representation-based Models for Text Matching through Virtual Interaction
Dan Li, Yang Yang, Hongyin Tang +4
With the booming of pre-trained transformers, representation-based models based on Siamese transformer encoders have become mainstream techniques for efficient text matching. Howev…
ASAP: A Chinese Review Dataset Towards Aspect Category Sentiment Analysis and Rating Prediction
Jiahao Bu, Lei Ren, Shuang Zheng +4
Sentiment analysis has attracted increasing attention in e-commerce. The sentiment polarities underlying user reviews are of great value for business intelligence. Aspect category…