activity
20152026
most citedBeyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

565 citations · 1k across the 38 of their papers we have counts for

collaborators
Showing 2019Show all

6 papers · 1 filter

cs.CL2019

QASC: A Dataset for Question Answering via Sentence Composition

Tushar Khot, Peter Clark, Michal Guerquin +2

Composing knowledge from multiple pieces of texts is a key challenge in multi-hop question answering. We present a multi-hop reasoning dataset, Question Answering via Sentence Comp…

cs.CL2019

What's Missing: A Knowledge Gap Guided Approach for Multi-hop Question Answering

Tushar Khot, Ashish Sabharwal, Peter Clark

Multi-hop textual question answering requires combining information from multiple sentences. We focus on a natural setting where, unlike typical reading comprehension, only partial…

cs.CL2019

From 'F' to 'A' on the N.Y. Regents Science Exams: An Overview of the Aristo Project

Peter Clark, Oren Etzioni, Daniel Khashabi +11

AI has achieved remarkable mastery over games such as Chess, Go, and Poker, and even Jeopardy, but the rich variety of standardized exams has remained a landmark challenge. Even in…

cs.CL2019★ 56 cited

Question Answering as Global Reasoning over Semantic Abstractions

Daniel Khashabi, Tushar Khot, Ashish Sabharwal +1

We propose a novel method for exploiting the semantic structure of text to answer multiple-choice questions. The approach is especially suitable for domains that require reasoning…

cs.CL2019★ 3 cited

Repurposing Entailment for Multi-Hop Question Answering Tasks

Harsh Trivedi, Heeyoung Kwon, Tushar Khot +2

Question Answering (QA) naturally reduces to an entailment problem, namely, verifying whether some text entails the answer to a question. However, for multi-hop QA tasks, which req…

cs.CL2019

On the Possibilities and Limitations of Multi-hop Reasoning Under Linguistic Imperfections

Daniel Khashabi, Erfan Sadeqi Azer, Tushar Khot +2

Systems for language understanding have become remarkably strong at overcoming linguistic imperfections in tasks involving phrase matching or simple reasoning. Yet, their accuracy…