15 citations · 31 across the 9 of their papers we have counts for
9 papers · 1 filter
Improving Large-scale Paraphrase Acquisition and Generation
Yao Dou, Chao Jiang, Wei Xu
This paper addresses the quality issues in existing Twitter-based paraphrase datasets, and discusses the necessity of using two separate definitions of paraphrase for identificatio…
arXivEdits: Understanding the Human Revision Process in Scientific Writing
Chao Jiang, Wei Xu, Samuel Stevens
Scientific publications are the primary means to communicate research discoveries, where the writing quality is of crucial importance. However, prior work studying the human editin…
Stanceosaurus: Classifying Stance Towards Multilingual Misinformation
Jonathan Zheng, Ashutosh Baheti, Tarek Naous +2
We present Stanceosaurus, a new corpus of 28,033 tweets in English, Hindi, and Arabic annotated with stance towards 251 misinformation claims. As far as we are aware, it is the lar…
BiSECT: Learning to Split and Rephrase Sentences with Bitexts
Joongwon Kim, Mounica Maddela, Reno Kriz +2
An important task in NLP applications such as sentence simplification is the ability to take a long, complex sentence and split it into shorter sentences, rephrasing as necessary.…
Neural semi-Markov CRF for Monolingual Word Alignment
Wuwei Lan, Chao Jiang, Wei Xu
Monolingual word alignment is important for studying fine-grained editing operations (i.e., deletion, addition, and substitution) in text-to-text generation tasks, such as paraphra…
Neural Network Models for Paraphrase Identification, Semantic Textual Similarity, Natural Language Inference, and Question Answering
Wuwei Lan, Wei Xu
In this paper, we analyze several neural network designs (and their variations) for sentence pair modeling and compare their performance extensively across eight datasets, includin…