14 citations · 35 across the 5 of their papers we have counts for
4 papers · 1 filter
Magic Pyramid: Accelerating Inference with Early Exiting and Token Pruning
Xuanli He, Iman Keivanloo, Yi Xu +4
Pre-training and then fine-tuning large language models is commonly used to achieve state-of-the-art performance in natural language processing (NLP) tasks. However, most pre-train…
Dialogue-oriented Pre-training
Yi Xu, Hai Zhao
Pre-trained language models (PrLM) has been shown powerful in enhancing a broad range of downstream tasks including various dialogue related ones. However, PrLMs are usually traine…
Topic-Aware Multi-turn Dialogue Modeling
Yi Xu, Hai Zhao, Zhuosheng Zhang
In the retrieval-based multi-turn dialogue modeling, it remains a challenge to select the most appropriate response according to extracting salient features in context utterances.…
On Leveraging the Visual Modality for Neural Machine Translation
Vikas Raunak, Sang Keun Choe, Quanyang Lu +2
Leveraging the visual modality effectively for Neural Machine Translation (NMT) remains an open problem in computational linguistics. Recently, Caglayan et al. posit that the obser…