416 citations · 615 across the 39 of their papers we have counts for
49 papers · 1 filter
Soft-Labeled Contrastive Pre-training for Function-level Code Representation
Xiaonan Li, Daya Guo, Yeyun Gong +6
Code contrastive pre-training has recently achieved significant progress on code-related tasks. In this paper, we present \textbf{SCodeR}, a \textbf{S}oft-labeled contrastive pre-t…
Mixed-modality Representation Learning and Pre-training for Joint Table-and-Text Retrieval in OpenQA
Junjie Huang, Wanjun Zhong, Qian Liu +3
Retrieving evidences from tabular and textual resources is essential for open-domain question answering (OpenQA), which provides more comprehensive information. However, training a…
Bridging the Gap between Language Models and Cross-Lingual Sequence Labeling
Nuo Chen, Linjun Shou, Ming Gong +2
Large-scale cross-lingual pre-trained language models (xPLMs) have shown effectiveness in cross-lingual sequence labeling tasks (xSL), such as cross-lingual machine reading compreh…
Multi-View Document Representation Learning for Open-Domain Dense Retrieval
Shunyu Zhang, Yaobo Liang, Ming Gong +2
Dense retrieval has achieved impressive advances in first-stage retrieval from a large-scale document collection, which is built on bi-encoder architecture to produce single vector…
EventBERT: A Pre-Trained Model for Event Correlation Reasoning
Yucheng Zhou, Xiubo Geng, Tao Shen +2
Event correlation reasoning infers whether a natural language paragraph containing multiple events conforms to human common sense. For example, "Andrew was very drowsy, so he took…
Building an Efficient and Effective Retrieval-based Dialogue System via Mutual Learning
Chongyang Tao, Jiazhan Feng, Chang Liu +3
Establishing retrieval-based dialogue systems that can select appropriate responses from the pre-built index has gained increasing attention from researchers. For this task, the ad…