activity
20172022
most citedUnified Pretraining Framework for Document Understanding

17 citations · 20 across the 7 of their papers we have counts for

collaborators

9 papers

cs.CL2021

From Toxicity in Online Comments to Incivility in American News: Proceed with Caution

Anushree Hede, Oshin Agarwal, Linda Lu +2

The ability to quantify incivility online, in news and in congressional debates, is of great interest to political scientists. Computational tools for detecting online incivility f…

cs.IR2020

Trialstreamer: Mapping and Browsing Medical Evidence in Real-Time

Benjamin E. Nye, Ani Nenkova, Iain J. Marshall +1

We introduce Trialstreamer, a living database of clinical trial reports. Here we mainly describe the evidence extraction component; this extracts from biomedical abstracts key piec…

cs.CL2020

Interpretability Analysis for Named Entity Recognition to Understand System Predictions and How They Can Improve

Oshin Agarwal, Yinfei Yang, Byron C. Wallace +1

Named Entity Recognition systems achieve remarkable performance on domains such as English news. It is natural to ask: What are these models actually learning to achieve this? Are…

cs.CL2020

Entity-Switched Datasets: An Approach to Auditing the In-Domain Robustness of Named Entity Recognition Models

Oshin Agarwal, Yinfei Yang, Byron C. Wallace +1

Named entity recognition systems perform well on standard datasets comprising English news. But given the paucity of data, it is difficult to draw conclusions about the robustness…

cs.CL2019

Predicting Annotation Difficulty to Improve Task Routing and Model Performance for Biomedical Information Extraction

Yinfei Yang, Oshin Agarwal, Chris Tar +2

Modern NLP systems require high-quality annotated data. In specialized domains, expert annotations may be prohibitively expensive. An alternative is to rely on crowdsourcing to red…

cs.CL2018

Named Person Coreference in English News

Oshin Agarwal, Sanjay Subramanian, Ani Nenkova +1

People are often entities of interest in tasks such as search and information extraction. In these tasks, the goal is to find as much information as possible about people specified…