activity
20182022
most citedDynatask: A Framework for Creating Dynamic AI Benchmark Tasks

2 citations · 2 across the 1 of their papers we have counts for

collaborators

5 papers

cs.CL20222 cited

Dynatask: A Framework for Creating Dynamic AI Benchmark Tasks

Tristan Thrush, Kushal Tirumala, Anmol Gupta +7

We introduce Dynatask: an open source system for setting up custom NLP tasks that aims to greatly lower the technical knowledge and effort required for hosting and evaluating state…

cs.CL2020

Information Seeking in the Spirit of Learning: a Dataset for Conversational Curiosity

Pedro Rodriguez, Paul Crook, Seungwhan Moon +1

Open-ended human learning and information-seeking are increasingly mediated by digital assistants. However, such systems often ignore the user's pre-existing knowledge. Assuming a…

cs.CL2019

Mitigating Noisy Inputs for Question Answering

Denis Peskov, Joe Barrow, Pedro Rodriguez +2

Natural language processing systems are often downstream of unreliable inputs: machine translation, optical character recognition, or speech recognition. For instance, virtual assi…

cs.CL2019

Quizbowl: The Case for Incremental Question Answering

Pedro Rodriguez, Shi Feng, Mohit Iyyer +2

Scholastic trivia competitions test knowledge and intelligence through mastery of question answering. Modern question answering benchmarks are one variant of the Turing test. Speci…

cs.CL2018

Trick Me If You Can: Human-in-the-loop Generation of Adversarial Examples for Question Answering

Eric Wallace, Pedro Rodriguez, Shi Feng +2

Adversarial evaluation stress tests a model's understanding of natural language. While past approaches expose superficial patterns, the resulting adversarial examples are limited i…