SearchQA: A New Q&A Dataset Augmented with Context from a Search Engine
arXiv:1704.05179
Abstract
We publicly release a new large-scale dataset, called SearchQA, for machine comprehension, or question-answering. Unlike recently released datasets, such as DeepMind CNN/DailyMail and SQuAD, the proposed SearchQA was constructed to reflect a full pipeline of general question-answering. That is, we start not from an existing article and generate a question-answer pair, but start from an existing question-answer pair, crawled from J! Archive, and augment it with text snippets retrieved by Google. Following this approach, we built SearchQA, which consists of more than 140k question-answer pairs with each pair having 49.6 snippets on average. Each question-answer-context tuple of the SearchQA comes with additional meta-data such as the snippet's URL, which we believe will be valuable resources for future research. We conduct human evaluation as well as test two baseline methods, one simple word selection and the other deep learning based, on the SearchQA. We show that there is a meaningful gap between the human and machine performances. This suggests that the proposed dataset could well serve as a benchmark for question-answering.
References in corpus (3)
Cited by in corpus (121)
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
- CTRL: A Conditional Transformer Language Model for Controllable Generation
- BERT Post-Training for Review Reading Comprehension and Aspect-based Sentiment Analysis
- Passage Re-ranking with BERT
- ReCoRD: Bridging the Gap between Human and Machine Commonsense Reading Comprehension
- Measuring Robustness to Natural Distribution Shifts in Image Classification
- QA Dataset Explosion: A Taxonomy of NLP Resources for Question Answering and Reading Comprehension
- Evidence Aggregation for Answer Re-Ranking in Open-Domain Question Answering
- Retrieving and Reading: A Comprehensive Survey on Open-domain Question Answering
- K-Adapter: Infusing Knowledge into Pre-Trained Models with Adapters
- Semantic Models for the First-stage Retrieval: A Comprehensive Review
- HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
- Rethinking Search: Making Domain Experts out of Dilettantes
- R: Reinforced Reader-Ranker for Open-Domain Question Answering
- Exploring Graph-structured Passage Representation for Multi-hop Reading Comprehension with Graph Neural Networks
- A Joint Model for Question Answering and Question Generation
- Open-Retrieval Conversational Question Answering
- What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams
- DuReader: a Chinese Machine Reading Comprehension Dataset from Real-world Applications
- Machine Reading Comprehension: The Role of Contextualized Language Models and Beyond
- Entity-Relation Extraction as Multi-Turn Question Answering
- Interactive Question Answering Systems: Literature Review
- Universal Text Representation from BERT: An Empirical Study
- A Dataset for Answering Time-Sensitive Questions
- PullNet: Open Domain Question Answering with Iterative Retrieval on Knowledge Bases and Text
- AQuaMuSe: Automatically Generating Datasets for Query-Based Multi-Document Summarization
- Enhancing lexical-based approach with external knowledge for Vietnamese multiple-choice machine reading comprehension
- Multi-range Reasoning for Machine Comprehension
- Retrieve-and-Read: Multi-task Learning of Information Retrieval and Reading Comprehension
- TyDi QA: A Benchmark for Information-Seeking Question Answering in Typologically Diverse Languages
- Improving Question Answering with External Knowledge
- Neural Machine Reading Comprehension: Methods and Trends
- Efficient and Robust Question Answering from Minimal Context over Documents
- Knowledge Enhanced Pretrained Language Models: A Compreshensive Survey
- Pretrained Encyclopedia: Weakly Supervised Knowledge-Pretrained Language Model
- Focused Hierarchical RNNs for Conditional Sequence Processing
- Multi-passage BERT: A Globally Normalized BERT Model for Open-domain Question Answering
- A Survey of Knowledge Enhanced Pre-trained Models
- The Effect of Natural Distribution Shift on Question Answering Models
- Contextualized Representations Using Textual Encyclopedic Knowledge
- MRQA 2019 Shared Task: Evaluating Generalization in Reading Comprehension
- Question Answering is a Format; When is it Useful?
- Coreferential Reasoning Learning for Language Representation
- Cluster-Former: Clustering-based Sparse Transformer for Long-Range Dependency Encoding
- Beyond Leaderboards: A survey of methods for revealing weaknesses in Natural Language Inference data and models
- Weaver: Deep Co-Encoding of Questions and Documents for Machine Reading
- Review Conversational Reading Comprehension
- PQuAD: A Persian Question Answering Dataset
- Blockwise Self-Attention for Long Document Understanding
- AmbigQA: Answering Ambiguous Open-domain Questions
- MultiReQA: A Cross-Domain Evaluation for Retrieval Question Answering Models
- Towards Domain Adaptation from Limited Data for Question Answering Using Deep Neural Networks
- ELI5: Long Form Question Answering
- Multi-Paragraph Reasoning with Knowledge-enhanced Graph Neural Network
- Probabilistic Assumptions Matter: Improved Models for Distantly-Supervised Document-Level Question Answering
- MMM: Multi-stage Multi-task Learning for Multi-choice Reading Comprehension
- Answering Science Exam Questions Using Query Rewriting with Background Knowledge
- HybridQA: A Dataset of Multi-Hop Question Answering over Tabular and Textual Data
- Multi-Passage Machine Reading Comprehension with Cross-Passage Answer Verification
- Interactive Language Learning by Question Answering
- Joint Training of Candidate Extraction and Answer Selection for Reading Comprehension
- Understanding Questions that Arise When Working with Business Documents
- Knowledge-Aided Open-Domain Question Answering
- Multi-Stage Conversational Passage Retrieval: An Approach to Fusing Term Importance Estimation and Neural Query Rewriting
- A Survey on Machine Reading Comprehension: Tasks, Evaluation Metrics and Benchmark Datasets
- CAiRE-COVID: A Question Answering and Query-focused Multi-Document Summarization System for COVID-19 Scholarly Information Management
- More Than Reading Comprehension: A Survey on Datasets and Metrics of Textual Question Answering
- Evidence Sentence Extraction for Machine Reading Comprehension
- Controlling Risk of Web Question Answering
- Cosmos QA: Machine Reading Comprehension with Contextual Commonsense Reasoning
- Learning to Coordinate Multiple Reinforcement Learning Agents for Diverse Query Reformulation
- Perhaps PTLMs Should Go to School -- A Task to Assess Open Book and Closed Book QA
- Building Efficient and Effective OpenQA Systems for Low-Resource Languages
- Quizbowl: The Case for Incremental Question Answering
- Towards Automatic Generation of Questions from Long Answers
- HAS-QA: Hierarchical Answer Spans Model for Open-domain Question Answering
- Bayesian Multi-Task Transfer Learning for Soft Prompt Tuning
- BIOMRC: A Dataset for Biomedical Machine Reading Comprehension
- Selective Question Answering under Domain Shift
- Multi-span Style Extraction for Generative Reading Comprehension
- Retrieve, Read, Rerank: Towards End-to-End Multi-Document Reading Comprehension
- ReCO: A Large Scale Chinese Reading Comprehension Dataset on Opinion
- FastFusionNet: New State-of-the-Art for DAWNBench SQuAD
- AmazonQA: A Review-Based Question Answering Task
- DREAM: A Challenge Dataset and Models for Dialogue-Based Reading Comprehension
- Analyzing Language Learned by an Active Question Answering Agent
- MATINF: A Jointly Labeled Large-Scale Dataset for Classification, Question Answering and Summarization
- Phrase-Indexed Question Answering: A New Challenge for Scalable Document Comprehension
- Composing Answer from Multi-spans for Reading Comprehension
- Answering Open-Domain Questions of Varying Reasoning Steps from Text
- Exploring Fluent Query Reformulations with Text-to-Text Transformers and Reinforcement Learning
- CrossFit: A Few-shot Learning Challenge for Cross-task Generalization in NLP
- Less is More: Rejecting Unreliable Reviews for Product Question Answering
- InfoTech Assistant: A Multimodal Conversational Agent for InfoTechnology Web Portal Queries
- Large Scale Question Answering using Tourism Data
- XCMRC: Evaluating Cross-lingual Machine Reading Comprehension
- ManyModalQA: Modality Disambiguation and QA over Diverse Inputs
- Zero-Shot Dialogue State Tracking via Cross-Task Transfer
- Weakly-Supervised Open-Retrieval Conversational Question Answering
- Tackling Graphical NLP problems with Graph Recurrent Networks
- LiveQA: A Question Answering Dataset over Sports Live
- How Optimal is Greedy Decoding for Extractive Question Answering?
- Knowledge Efficient Deep Learning for Natural Language Processing
- Syntax-Enhanced Pre-trained Model
- A Study of the Tasks and Models in Machine Reading Comprehension
- Distantly-Supervised Evidence Retrieval Enables Question Answering without Evidence Annotation
- Narrative Question Answering with Cutting-Edge Open-Domain QA Techniques: A Comprehensive Study
- Enhance Long Text Understanding via Distilled Gist Detector from Abstractive Summarization
- Towards an Atlas of Cultural Commonsense for Machine Reasoning
- Improve Query Focused Abstractive Summarization by Incorporating Answer Relevance
- When to Fold'em: How to answer Unanswerable questions
- Generative Context Pair Selection for Multi-hop Question Answering
- An Exploration of Data Augmentation and Sampling Techniques for Domain-Agnostic Question Answering
- Knowing More About Questions Can Help: Improving Calibration in Question Answering
- CAiRE in DialDoc21: Data Augmentation for Information-Seeking Dialogue System
- Contrastive Domain Adaptation for Question Answering using Limited Text Corpora
- Inquisitive Question Generation for High Level Text Comprehension
- Multi-Relational Question Answering from Narratives: Machine Reading and Reasoning in Simulated Worlds
- RoR: Read-over-Read for Long Document Machine Reading Comprehension
- WebSRC: A Dataset for Web-Based Structural Reading Comprehension
- SRLGRN: Semantic Role Labeling Graph Reasoning Network