A large annotated corpus for learning natural language inference
arXiv:1508.05326
Abstract
Understanding entailment and contradiction is fundamental to understanding natural language, and inference about entailment and contradiction is a valuable testing ground for the development of semantic representations. However, machine learning research in this area has been dramatically limited by the lack of large-scale resources. To address this, we introduce the Stanford Natural Language Inference corpus, a new, freely available collection of labeled sentence pairs, written by humans doing a novel grounded task based on image captioning. At 570K pairs, it is two orders of magnitude larger than all other resources of its type. This increase in scale allows lexicalized classifiers to outperform some sophisticated existing entailment models, and it allows a neural network-based model to perform competitively on natural language inference benchmarks for the first time.
To appear at EMNLP 2015. The data will be posted shortly before the conference (the week of 14 Sep) at http://nlp.stanford.edu/projects/snli/
References in corpus (1)
Cited by in corpus (260)
- Cross-lingual Language Model Pretraining
- Universal Sentence Encoder
- Recent Trends in Deep Learning Based Natural Language Processing
- Enhanced LSTM for Natural Language Inference
- Comparative Study of CNN and RNN for Natural Language Processing
- Align before Fuse: Vision and Language Representation Learning with Momentum Distillation
- Reasoning about Entailment with Neural Attention
- Generating Natural Adversarial Examples
- From Softmax to Sparsemax: A Sparse Model of Attention and Multi-Label Classification
- Deep Learning Based Text Classification: A Comprehensive Review
- SentEval: An Evaluation Toolkit for Universal Sentence Representations
- The Natural Language Decathlon: Multitask Learning as Question Answering
- Pathologies of Neural Models Make Interpretations Difficult
- TabFact: A Large-scale Dataset for Table-based Fact Verification
- Learning Natural Language Inference using Bidirectional LSTM model and Inner-Attention
- CLEAR: Contrastive Learning for Sentence Representation
- ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation
- You Only Look at One Sequence: Rethinking Transformer in Vision through Object Detection
- Long Short-Term Memory-Networks for Machine Reading
- Supervised Multimodal Bitransformers for Classifying Images and Text
- Visual Entailment: A Novel Task for Fine-Grained Image Understanding
- XNLI: Evaluating Cross-lingual Sentence Representations
- A Survey of Word Embeddings Evaluation Methods
- Transforming Question Answering Datasets Into Natural Language Inference Datasets
- Bilateral Multi-Perspective Matching for Natural Language Sentences
- Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training
- DiSAN: Directional Self-Attention Network for RNN/CNN-Free Language Understanding
- Neural Network Models for Paraphrase Identification, Semantic Textual Similarity, Natural Language Inference, and Question Answering
- Pretrained Transformers for Text Ranking: BERT and Beyond
- Entailment as Few-Shot Learner
- WT5?! Training Text-to-Text Models to Explain their Predictions
- Evaluation of sentence embeddings in downstream and linguistic probing tasks
- A Survey on Text Classification: From Shallow to Deep Learning
- StructBERT: Incorporating Language Structures into Pre-training for Deep Language Understanding
- Learning General Purpose Distributed Sentence Representations via Large Scale Multi-task Learning
- Is BERT Really Robust? A Strong Baseline for Natural Language Attack on Text Classification and Entailment
- Beyond 512 Tokens: Siamese Multi-depth Transformer-based Hierarchical Encoder for Long-Form Document Matching
- Deep Enhanced Representation for Implicit Discourse Relation Recognition
- Order-Embeddings of Images and Language
- How Transferable are Neural Networks in NLP Applications?
- Jointly Learning Sentence Embeddings and Syntax with Unsupervised Tree-LSTMs
- Zero-Shot Cross-lingual Classification Using Multilingual Neural Machine Translation
- Probabilistic Reasoning via Deep Learning: Neural Association Models
- Sparse Sinkhorn Attention
- Reconstruction of turbulent data with deep generative models for semantic inpainting from TURB-Rot database
- ERNIE 2.0: A Continual Pre-training Framework for Language Understanding
- Graph Neural Networks for Natural Language Processing: A Survey
- Align, Mask and Select: A Simple Method for Incorporating Commonsense Knowledge into Language Representation Models
- Generating Persona Consistent Dialogues by Exploiting Natural Language Inference
- Distance-based Self-Attention Network for Natural Language Inference
- BERT-ATTACK: Adversarial Attack Against BERT Using BERT
- Adversarial NLI: A New Benchmark for Natural Language Understanding
- Learning to Compute Word Embeddings On the Fly
- Learning Natural Language Inference with LSTM
- Dynamic Integration of Background Knowledge in Neural NLU Systems
- A Decomposable Attention Model for Natural Language Inference
- Understanding and Mitigating the Security Risks of Voice-Controlled Third-Party Skills on Amazon Alexa and Google Home
- A Survey on Transfer Learning in Natural Language Processing
- Learning from others' mistakes: Avoiding dataset biases without modeling them
- Adversarial Examples in Modern Machine Learning: A Review
- Natural Language Inference by Tree-Based Convolution and Heuristic Matching
- Learning to Compose Words into Sentences with Reinforcement Learning
- Reweighted Proximal Pruning for Large-Scale Language Representation
- An Objective Metric for Explainable AI: How and Why to Estimate the Degree of Explainability
- Adversarially Regularising Neural NLI Models to Integrate Logical Background Knowledge
- Charagram: Embedding Words and Sentences via Character n-grams
- Semantic Sentence Matching with Densely-connected Recurrent and Co-attentive Information
- Adversarially Regularized Autoencoders
- Shortcut-Stacked Sentence Encoders for Multi-Domain Inference
- Can I Trust the Explainer? Verifying Post-hoc Explanatory Methods
- Hyperbolic Neural Networks
- Dynamic Neural Turing Machine with Soft and Hard Addressing Schemes
- Learning Gender-Neutral Word Embeddings
- A Survey on Explainability in Machine Reading Comprehension
- Learning Semantic Textual Similarity from Conversations
- Are we pretraining it right? Digging deeper into visio-linguistic pretraining
- Combining Fact Extraction and Verification with Neural Semantic Matching Networks
- Probing Biomedical Embeddings from Language Models
- Towards a Robust Deep Neural Network in Texts: A Survey
- Reinforced Self-Attention Network: a Hybrid of Hard and Soft Attention for Sequence Modeling
- Sentence Pair Scoring: Towards Unified Framework for Text Comprehension
- Uncertainty-Aware Reliable Text Classification
- On the Effectiveness of Low-Rank Matrix Factorization for LSTM Model Compression
- Interpreting Deep Learning Models in Natural Language Processing: A Review
- Self-Explaining Structures Improve NLP Models
- Fake News Detection as Natural Language Inference
- Unsupervised Topic Segmentation of Meetings with BERT Embeddings
- KG^2: Learning to Reason Science Exam Questions with Contextual Knowledge Graph Embeddings
- A New Dataset for Natural Language Inference from Code-mixed Conversations
- Adversarial attacks against Fact Extraction and VERification
- ACTRCE: Augmenting Experience via Teacher's Advice For Multi-Goal Reinforcement Learning
- Read + Verify: Machine Reading Comprehension with Unanswerable Questions
- What's in the Box? A Preliminary Analysis of Undesirable Content in the Common Crawl Corpus
- Bilingual Embeddings with Random Walks over Multilingual Wordnets
- Lessons from Natural Language Inference in the Clinical Domain
- Lightweight and Efficient Neural Natural Language Processing with Quaternion Networks
- Back to the Future: Unsupervised Backprop-based Decoding for Counterfactual and Abductive Commonsense Reasoning
- Learning beyond datasets: Knowledge Graph Augmented Neural Networks for Natural language Processing
- Beyond Leaderboards: A survey of methods for revealing weaknesses in Natural Language Inference data and models
- HUBERT Untangles BERT to Improve Transfer across NLP Tasks
- Improving Natural Language Inference Using External Knowledge in the Science Questions Domain
- Modelling Sentence Pairs with Tree-structured Attentive Encoder
- Learning semantic sentence representations from visually grounded language without lexical knowledge
- Generating Natural Language Inference Chains
- LIREx: Augmenting Language Inference with Relevant Explanation
- Visual Entailment Task for Visually-Grounded Language Learning
- Answering Science Exam Questions Using Query Rewriting with Background Knowledge
- Parallelizing Legendre Memory Unit Training
- The Importance of Being Recurrent for Modeling Hierarchical Structure
- Cross-Language Learning for Program Classification using Bilateral Tree-Based Convolutional Neural Networks
- Modelling Interaction of Sentence Pair with coupled-LSTMs
- Transfer Learning for Context-Aware Question Matching in Information-seeking Conversations in E-commerce
- Contradiction Detection for Rumorous Claims
- Story Ending Prediction by Transferable BERT
- Knowledge Enhanced Attention for Robust Natural Language Inference
- UKP-Athene: Multi-Sentence Textual Entailment for Claim Verification
- Semi Supervised Preposition-Sense Disambiguation using Multilingual Data
- Can Neural Networks Understand Logical Entailment?
- Interpreting Recurrent and Attention-Based Neural Models: a Case Study on Natural Language Inference
- Self-Guided Contrastive Learning for BERT Sentence Representations
- Recurrent Neural Network Encoder with Attention for Community Question Answering
- Natural Language Generation with Neural Variational Models
- Multimodal Fusion Refiner Networks
- Neural Paraphrase Identification of Questions with Noisy Pretraining
- Dynamic Compositional Neural Networks over Tree Structure
- Improving Classification through Weak Supervision in Context-specific Conversational Agent Development for Teacher Education
- The Benchmark Lottery
- Analyzing Compositionality-Sensitivity of NLI Models
- A Qualitative Comparison of CoQA, SQuAD 2.0 and QuAC
- Dynamic Multi-Level Multi-Task Learning for Sentence Simplification
- Learning to update Auto-associative Memory in Recurrent Neural Networks for Improving Sequence Memorization
- Learning Structured Text Representations
- A Neural Architecture Mimicking Humans End-to-End for Natural Language Inference
- On the Effective Use of Pretraining for Natural Language Inference
- Story Ending Generation with Incremental Encoding and Commonsense Knowledge
- Effective writing style imitation via combinatorial paraphrasing
- Natural Language Inference over Interaction Space: ICLR 2018 Reproducibility Report
- Multi-task Learning for Universal Sentence Embeddings: A Thorough Evaluation using Transfer and Auxiliary Tasks
- Order in the Court: Explainable AI Methods Prone to Disagreement
- Deep Learning Based on Generative Adversarial and Convolutional Neural Networks for Financial Time Series Predictions
- Learning to Reason With Adaptive Computation
- A Hybrid Approach to Measure Semantic Relatedness in Biomedical Concepts
- Semantics Preserving Adversarial Learning
- Scheduled DropHead: A Regularization Method for Transformer Models
- Counterfactual Variable Control for Robust and Interpretable Question Answering
- Marked Attribute Bias in Natural Language Inference
- DELTA: A DEep learning based Language Technology plAtform
- Exploring Neural Models for Parsing Natural Language into First-Order Logic
- Joint Multi-Domain Learning for Automatic Short Answer Grading
- Training a Ranking Function for Open-Domain Question Answering
- Contextual Text Denoising with Masked Language Models
- AR-LSAT: Investigating Analytical Reasoning of Text
- Poison Attacks against Text Datasets with Conditional Adversarially Regularized Autoencoder
- Identifying Well-formed Natural Language Questions
- A Generate-Validate Approach to Answering Questions about Qualitative Relationships
- Sieving Fake News From Genuine: A Synopsis
- A Bayesian Approach to Direct and Inverse Abstract Argumentation Problems
- Parameter Re-Initialization through Cyclical Batch Size Schedules
- Neural Skill Transfer from Supervised Language Tasks to Reading Comprehension
- Multi-view Sentence Representation Learning
- MVP-BERT: Redesigning Vocabularies for Chinese BERT and Multi-Vocab Pretraining
- Second-Order Word Embeddings from Nearest Neighbor Topological Features
- Pre-training for Abstractive Document Summarization by Reinstating Source Text
- An Empirical Evaluation of various Deep Learning Architectures for Bi-Sequence Classification Tasks
- ReCO: A Large Scale Chinese Reading Comprehension Dataset on Opinion
- TextNAS: A Neural Architecture Search Space tailored for Text Representation
- Augmenting Neural Networks with First-order Logic
- Path-Based Contextualization of Knowledge Graphs for Textual Entailment
- : Author Attribute Anonymity by Adversarial Training of Neural Machine Translation
- Detecting and Explaining Causes From Text For a Time Series Event
- Learning Robust, Transferable Sentence Representations for Text Classification
- Attention Boosted Sequential Inference Model
- Sentence Encoding with Tree-constrained Relation Networks
- Improving Sentence Representations with Consensus Maximisation
- What If We Simply Swap the Two Text Fragments? A Straightforward yet Effective Way to Test the Robustness of Methods to Confounding Signals in Nature Language Inference Tasks
- CMV-BERT: Contrastive multi-vocab pretraining of BERT
- Rethinking Skip-thought: A Neighborhood based Approach
- Evaluation of Unsupervised Compositional Representations
- Inducing Grammars with and for Neural Machine Translation
- Mimic and Conquer: Heterogeneous Tree Structure Distillation for Syntactic NLP
- Direct Network Transfer: Transfer Learning of Sentence Embeddings for Semantic Similarity
- Phrase-Indexed Question Answering: A New Challenge for Scalable Document Comprehension
- Automatic Validation of Textual Attribute Values in E-commerce Catalog by Learning with Limited Labeled Data
- Attentive Convolution: Equipping CNNs with RNN-style Attention Mechanisms
- Modelling Domain Relationships for Transfer Learning on Retrieval-based Question Answering Systems in E-commerce
- Incorporating Domain Knowledge To Improve Topic Segmentation Of Long MOOC Lecture Videos
- Syntax-based Attention Model for Natural Language Inference
- Building an Evaluation Scale using Item Response Theory
- Towards Variable-Length Textual Adversarial Attacks
- Combining Axiom Injection and Knowledge Base Completion for Efficient Natural Language Inference
- Evaluating Multimodal Representations on Visual Semantic Textual Similarity
- Reasoning over Vision and Language: Exploring the Benefits of Supplemental Knowledge
- Linguistically-Enriched and Context-Aware Zero-shot Slot Filling
- Sarcasm Analysis using Conversation Context
- Semantic Matching Against a Corpus: New Applications and Methods
- Summarizing Utterances from Japanese Assembly Minutes using Political Sentence-BERT-based Method for QA Lab-PoliInfo-2 Task of NTCIR-15
- Probing the Natural Language Inference Task with Automated Reasoning Tools
- Human acceptability judgements for extractive sentence compression
- NLI Data Sanity Check: Assessing the Effect of Data Corruption on Model Performance
- Simplifying Neural Machine Translation with Addition-Subtraction Twin-Gated Recurrent Networks
- WikiRef: Wikilinks as a route to recommending appropriate references for scientific Wikipedia pages
- Dropping Networks for Transfer Learning
- Term Definitions Help Hypernymy Detection
- Picking Apart Story Salads
- Learning Universal Sentence Representations with Mean-Max Attention Autoencoder
- Predicting the Argumenthood of English Prepositional Phrases
- Data-Efficient Language-Supervised Zero-Shot Learning with Self-Distillation
- Constructing Explainable Opinion Graphs from Review
- On Tree-Based Neural Sentence Modeling
- Resource-Size matters: Improving Neural Named Entity Recognition with Optimized Large Corpora
- Adversarial Self-Supervised Data-Free Distillation for Text Classification
- Answer Ranking for Product-Related Questions via Multiple Semantic Relations Modeling
- Dynamic Compositionality in Recursive Neural Networks with Structure-aware Tag Representations
- Character-based Neural Networks for Sentence Pair Modeling
- REGMAPR - Text Matching Made Easy
- Recursive Top-Down Production for Sentence Generation with Latent Trees
- Lexicosyntactic Inference in Neural Models
- 15 Keypoints Is All You Need
- Geometry matters: Exploring language examples at the decision boundary
- Compositional Language Understanding with Text-based Relational Reasoning
- End-Task Oriented Textual Entailment via Deep Explorations of Inter-Sentence Interactions
- Transfer Reward Learning for Policy Gradient-Based Text Generation
- Empirical Evaluation of Multi-task Learning in Deep Neural Networks for Natural Language Processing
- Learning to Embed Sentences Using Attentive Recursive Trees
- Multi-task Sentence Encoding Model for Semantic Retrieval in Question Answering Systems
- Reference and Document Aware Semantic Evaluation Methods for Korean Language Summarization
- AWE: Asymmetric Word Embedding for Textual Entailment
- Acquisition of Phrase Correspondences using Natural Deduction Proofs
- Super Tickets in Pre-Trained Language Models: From Model Compression to Improving Generalization
- Efficient Purely Convolutional Text Encoding
- Evaluating Document Coherence Modelling
- An Empirical Study of Extrapolation in Text Generation with Scalar Control
- TransAug: Translate as Augmentation for Sentence Embeddings
- COSTRA 1.0: A Dataset of Complex Sentence Transformations
- FastTrees: Parallel Latent Tree-Induction for Faster Sequence Encoding
- e-QRAQ: A Multi-turn Reasoning Dataset and Simulator with Explanations
- Using Statistical and Semantic Models for Multi-Document Summarization
- SNU_IDS at SemEval-2018 Task 12: Sentence Encoder with Contextualized Vectors for Argument Reasoning Comprehension
- Metric for Automatic Machine Translation Evaluation based on Universal Sentence Representations
- Generating Contradictory, Neutral, and Entailing Sentences
- Counterfactual Maximum Likelihood Estimation for Training Deep Networks
- Bag-of-Vector Embeddings of Dependency Graphs for Semantic Induction
- Optimizing Open-Ended Crowdsourcing: The Next Frontier in Crowdsourced Data Management
- Towards Language Agnostic Universal Representations
- Teaching Syntax by Adversarial Distraction
- Multi-level Head-wise Match and Aggregation in Transformer for Textual Sequence Matching
- SufiSent - Universal Sentence Representations Using Suffix Encodings
- Incorporating Temporal Information in Entailment Graph Mining
- DCT: Dynamic Compressive Transformer for Modeling Unbounded Sequence
- Text Information Aggregation with Centrality Attention
- Multiple Structural Priors Guided Self Attention Network for Language Understanding
- Hierarchical RNN with Static Sentence-Level Attention for Text-Based Speaker Change Detection
- Understanding Deep Learning Performance through an Examination of Test Set Difficulty: A Psychometric Case Study
- SCDE: Sentence Cloze Dataset with High Quality Distractors From Examinations
- Ordinal Common-sense Inference
- Exploring Lexical Irregularities in Hypothesis-Only Models of Natural Language Inference
- Pre-trained Language Model Based Active Learning for Sentence Matching
- Latte-Mix: Measuring Sentence Semantic Similarity with Latent Categorical Mixtures
- A strong baseline for question relevancy ranking
- Improving Matching Models with Hierarchical Contextualized Representations for Multi-turn Response Selection