activity
20192021
most citedDROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

96 citations · 112 across the 4 of their papers we have counts for

collaborators

7 papers

cs.CL2021

Generative Context Pair Selection for Multi-hop Question Answering

Dheeru Dua, Cicero Nogueira dos Santos, Patrick Ng +4

Compositional reasoning tasks like multi-hop question answering, require making latent decisions to get the final answer, given a question. However, crowdsourced datasets often cap…

cs.CL2021

Learning with Instance Bundles for Reading Comprehension

Dheeru Dua, Pradeep Dasigi, Sameer Singh +1

When training most modern reading comprehension models, all the questions associated with a context are treated as being independent from each other. However, closely related quest…

cs.HC2020

Easy, Reproducible and Quality-Controlled Data Collection with Crowdaq

Qiang Ning, Hao Wu, Pradeep Dasigi +5

High-quality and large-scale data are key to success for AI systems. However, large-scale data annotation efforts are often confronted with a set of common challenges: (1) designin…

cs.CL2020

Evaluating Models' Local Decision Boundaries via Contrast Sets

Matt Gardner, Yoav Artzi, Victoria Basmova +23

Standard test sets for supervised learning evaluate in-distribution generalization. Unfortunately, when a dataset has systematic gaps (e.g., annotation artifacts), these evaluation…

cs.CL201914 cited

ORB: An Open Reading Benchmark for Comprehensive Evaluation of Machine Reading Comprehension

Dheeru Dua, Ananth Gottumukkala, Alon Talmor +2

Reading comprehension is one of the crucial tasks for furthering research in natural language understanding. A lot of diverse reading comprehension datasets have recently been intr…

cs.CL201996 cited

DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Dheeru Dua, Yizhong Wang, Pradeep Dasigi +3

Reading comprehension has recently seen rapid progress, with systems matching humans on the most popular datasets for the task. However, a large body of work has highlighted the br…