The Goldilocks Principle: Reading Children's Books with Explicit Memory Representations
arXiv:1511.02301
Abstract
We introduce a new test of how well language models capture meaning in children's books. Unlike standard language modelling benchmarks, it distinguishes the task of predicting syntactic function words from that of predicting lower-frequency words, which carry greater semantic content. We compare a range of state-of-the-art models, each with a different way of encoding what has been previously read. We show that models which store explicit representations of long-term contexts outperform state-of-the-art neural language models at predicting semantic content words, although this advantage is not observed for syntactic function words. Interestingly, we find that the amount of text encoded in a single memory representation is highly influential to the performance: there is a sweet-spot, not too big and not too small, between single words and full sentences that allows the most meaningful information in a text to be effectively retained and recalled. Further, the attention over such window-based memories can be trained effectively through self-supervision. We then assess the generality of this principle by applying it to the CNN QA benchmark, which involves identifying named entities in paraphrased summaries of news articles, and achieve state-of-the-art performance.
References in corpus (4)
Cited by in corpus (123)
- Matching Networks for One Shot Learning
- SQuAD: 100,000+ Questions for Machine Comprehension of Text
- Attention in Natural Language Processing
- Dynamic Coattention Networks For Question Answering
- QANet: Combining Local Convolution with Global Self-Attention for Reading Comprehension
- SearchQA: A New Q&A Dataset Augmented with Context from a Search Engine
- BERT Post-Training for Review Reading Comprehension and Aspect-based Sentiment Analysis
- ReasoNet: Learning to Stop Reading in Machine Comprehension
- Learning to Remember Rare Events
- ReCoRD: Bridging the Gap between Human and Machine Commonsense Reading Comprehension
- StoryBuddy: A Human-AI Collaborative Chatbot for Parent-Child Interactive Storytelling with Flexible Parental Involvement
- QA Dataset Explosion: A Taxonomy of NLP Resources for Question Answering and Reading Comprehension
- Retrieving and Reading: A Comprehensive Survey on Open-domain Question Answering
- A Thorough Examination of the CNN/Daily Mail Reading Comprehension Task
- A Neural Knowledge Language Model
- RACE: Large-scale ReAding Comprehension Dataset From Examinations
- Reading Wikipedia to Answer Open-Domain Questions
- Neural Models for Information Retrieval
- DRCD: a Chinese Machine Reading Comprehension Dataset
- Exploring Graph-structured Passage Representation for Multi-hop Reading Comprehension with Graph Neural Networks
- Dialog-based Language Learning
- Query-Reduction Networks for Question Answering
- Probabilistic Reasoning via Deep Learning: Neural Association Models
- Findings of the BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
- MCScript: A Novel Dataset for Assessing Machine Comprehension Using Script Knowledge
- MEMEN: Multi-layer Embedding with Memory Networks for Machine Comprehension
- Adversarial NLI: A New Benchmark for Natural Language Understanding
- Sharp Nearby, Fuzzy Far Away: How Neural Language Models Use Context
- Recent Advances in Natural Language Inference: A Survey of Benchmarks, Resources, and Approaches
- Evaluating Prerequisite Qualities for Learning End-to-End Dialog Systems
- Embracing data abundance: BookTest Dataset for Reading Comprehension
- Stochastic Answer Networks for Natural Language Inference
- DuReader: a Chinese Machine Reading Comprehension Dataset from Real-world Applications
- Compressive Transformers for Long-Range Sequence Modelling
- Machine Reading Comprehension: The Role of Contextualized Language Models and Beyond
- TVQA: Localized, Compositional Video Question Answering
- Cross-Lingual Machine Reading Comprehension
- Larger-Context Language Modelling
- Words or Characters? Fine-grained Gating for Reading Comprehension
- Making Neural QA as Simple as Possible but not Simpler
- End-to-End Answer Chunk Extraction and Ranking for Reading Comprehension
- Reinforced Mnemonic Reader for Machine Reading Comprehension
- Exploring Question Understanding and Adaptation in Neural-Network-Based Question Answering
- U-Net: Machine Reading Comprehension with Unanswerable Questions
- Dynamic Neural Turing Machine with Soft and Hard Addressing Schemes
- Stochastic Answer Networks for Machine Reading Comprehension
- Consensus Attention-based Neural Networks for Chinese Reading Comprehension
- Enhancing lexical-based approach with external knowledge for Vietnamese multiple-choice machine reading comprehension
- A Comparative Study of Word Embeddings for Reading Comprehension
- Envisioning Narrative Intelligence: A Creative Visual Storytelling Anthology
- Convolutional Spatial Attention Model for Reading Comprehension with Multiple-Choice Questions
- Neural Machine Reading Comprehension: Methods and Trends
- Machine Reading Comprehension: a Literature Review
- Cognitive Graph for Multi-Hop Reading Comprehension at Scale
- Explicit Contextual Semantics for Text Comprehension
- Simulating Action Dynamics with Neural Process Networks
- A Parallel-Hierarchical Model for Machine Comprehension on Sparse Data
- A Co-Matching Model for Multi-choice Reading Comprehension
- SG-Net: Syntax-Guided Machine Reading Comprehension
- Deep Reinforcement Learning in Computer Vision: A Comprehensive Survey
- Subword-augmented Embedding for Cloze Reading Comprehension
- QA4IE: A Question Answering based Framework for Information Extraction
- Discriminative Sentence Modeling for Story Ending Prediction
- Review Conversational Reading Comprehension
- Weaver: Deep Co-Encoding of Questions and Documents for Machine Reading
- PQuAD: A Persian Question Answering Dataset
- Smarnet: Teaching Machines to Read and Comprehend Like Human
- What Makes Reading Comprehension Questions Easier?
- Ruminating Reader: Reasoning with Gated Multi-Hop Attention
- Discourse-Aware Semantic Self-Attention for Narrative Reading Comprehension
- Commonsense Knowledge Enhanced Embeddings for Solving Pronoun Disambiguation Problems in Winograd Schema Challenge
- ExpMRC: Explainability Evaluation for Machine Reading Comprehension
- HFL-RC System at SemEval-2018 Task 11: Hybrid Multi-Aspects Model for Commonsense Reading Comprehension
- Multi-Passage Machine Reading Comprehension with Cross-Passage Answer Verification
- Reasoning over RDF Knowledge Bases using Deep Learning
- RecipeQA: A Challenge Dataset for Multimodal Comprehension of Cooking Recipes
- Joint Training of Candidate Extraction and Answer Selection for Reading Comprehension
- DramaQA: Character-Centered Video Story Understanding with Hierarchical QA
- Can Neural Networks Understand Logical Entailment?
- Dataset for the First Evaluation on Chinese Machine Reading Comprehension
- More Than Reading Comprehension: A Survey on Datasets and Metrics of Textual Question Answering
- Understanding Dataset Design Choices for Multi-hop Reasoning
- A Survey on Machine Reading Comprehension: Tasks, Evaluation Metrics and Benchmark Datasets
- Evidence Sentence Extraction for Machine Reading Comprehension
- Cosmos QA: Machine Reading Comprehension with Contextual Commonsense Reasoning
- Recurrent Entity Networks with Delayed Memory Update for Targeted Aspect-based Sentiment Analysis
- Read, Retrospect, Select: An MRC Framework to Short Text Entity Linking
- Two-Stage Synthesis Networks for Transfer Learning in Machine Comprehension
- Hierarchical Question Answering for Long Documents
- Building Efficient and Effective OpenQA Systems for Low-Resource Languages
- A Vietnamese Dataset for Evaluating Machine Reading Comprehension
- Conditional Generation and Snapshot Learning in Neural Dialogue Systems
- Complex QA and language models hybrid architectures, Survey
- ChID: A Large-scale Chinese IDiom Dataset for Cloze Test
- Knowledge Based Machine Reading Comprehension
- Multi-Mention Learning for Reading Comprehension with Neural Cascades
- Adaptations of ROUGE and BLEU to Better Evaluate Machine Reading Comprehension Task
- Phrase-Indexed Question Answering: A New Challenge for Scalable Document Comprehension
- Dialogue Graph Modeling for Conversational Machine Reading
- A Pipeline for Creative Visual Storytelling
- Large-scale Cloze Test Dataset Created by Teachers
- Fine-tuning Strategies for Domain Specific Question Answering under Low Annotation Budget Constraints
- Contextual embedding and model weighting by fusing domain knowledge on Biomedical Question Answering
- PreCo: A Large-scale Dataset in Preschool Vocabulary for Coreference Resolution
- Effective Subword Segmentation for Text Comprehension
- Contextual Recurrent Units for Cloze-style Reading Comprehension
- TWEETQA: A Social Media Focused Question Answering Dataset
- Tackling Graphical NLP problems with Graph Recurrent Networks
- Unsupervised Explanation Generation for Machine Reading Comprehension
- A Boo(n) for Evaluating Architecture Performance
- XCMRC: Evaluating Cross-lingual Machine Reading Comprehension
- Composing Answer from Multi-spans for Reading Comprehension
- Dependent Gated Reading for Cloze-Style Question Answering
- Implicit Argument Prediction as Reading Comprehension
- A Study of the Tasks and Models in Machine Reading Comprehension
- A BERT-based Dual Embedding Model for Chinese Idiom Prediction
- Effective Character-augmented Word Embedding for Machine Reading Comprehension
- Memory networks for consumer protection:unfairness exposed
- NE-Table: A Neural key-value table for Named Entities
- SCDE: Sentence Cloze Dataset with High Quality Distractors From Examinations
- A Coarse to Fine Question Answering System based on Reinforcement Learning
- ClueReader: Heterogeneous Graph Attention Network for Multi-hop Machine Reading Comprehension
- Multi-Perspective Context Aggregation for Semi-supervised Cloze-style Reading Comprehension