activity
20172021
most citedA Continuously Growing Dataset of Sentential Paraphrases

15 citations · 17 across the 4 of their papers we have counts for

collaborators

7 papers

cs.CL2021

BiSECT: Learning to Split and Rephrase Sentences with Bitexts

Joongwon Kim, Mounica Maddela, Reno Kriz +2

An important task in NLP applications such as sentence simplification is the ability to take a long, complex sentence and split it into shorter sentences, rephrasing as necessary.…

cs.IR20211 cited

Pre-training for Ad-hoc Retrieval: Hyperlink is Also You Need

Zhengyi Ma, Zhicheng Dou, Wei Xu +4

Designing pre-training objectives that more closely resemble the downstream tasks for pre-trained language models can lead to better performance at the fine-tuning stage, especiall…

cs.CL20211 cited

Neural semi-Markov CRF for Monolingual Word Alignment

Wuwei Lan, Chao Jiang, Wei Xu

Monolingual word alignment is important for studying fine-grained editing operations (i.e., deletion, addition, and substitution) in text-to-text generation tasks, such as paraphra…

cs.CL2018

Neural Network Models for Paraphrase Identification, Semantic Textual Similarity, Natural Language Inference, and Question Answering

Wuwei Lan, Wei Xu

In this paper, we analyze several neural network designs (and their variations) for sentence pair modeling and compare their performance extensively across eight datasets, includin…

cs.CL2018

Character-based Neural Networks for Sentence Pair Modeling

Wuwei Lan, Wei Xu

Sentence pair modeling is critical for many NLP tasks, such as paraphrase identification, semantic textual similarity, and natural language inference. Most state-of-the-art neural…

cs.CL2018

An Annotated Corpus for Machine Reading of Instructions in Wet Lab Protocols

Chaitanya Kulkarni, Wei Xu, Alan Ritter +1

We describe an effort to annotate a corpus of natural language instructions consisting of 622 wet lab protocols to facilitate automatic or semi-automatic conversion of protocols in…