Towards Explainable Fact Checking
arXiv:2108.10274
Abstract
The past decade has seen a substantial rise in the amount of mis- and disinformation online, from targeted disinformation campaigns to influence politics, to the unintentional spreading of misinformation about public health. This development has spurred research in the area of automatic fact checking, from approaches to detect check-worthy claims and determining the stance of tweets towards claims, to methods to determine the veracity of claims given evidence documents. These automatic methods are often content-based, using natural language processing methods, which in turn utilise deep neural networks to learn higher-order features from text in order to make predictions. As deep neural networks are black-box models, their inner workings cannot be easily explained. At the same time, it is desirable to explain how they arrive at certain decisions, especially if they are to be used for decision making. While this has been known for some time, the issues this raises have been exacerbated by models increasing in size, and by EU legislation requiring models to be used for decision making to provide explanations, and, very recently, by legislation requiring online platforms operating in the EU to provide transparent reporting on their services. Despite this, current solutions for explainability are still lacking in the area of fact checking. This thesis presents my research on automatic fact checking, including claim check-worthiness detection, stance detection and veracity prediction. Its contributions go beyond fact checking, with the thesis proposing more general machine learning solutions for natural language processing in the area of learning with limited labelled data. Finally, the thesis presents some first solutions for explainable fact checking.
Thesis presented to the University of Copenhagen Faculty of Science in partial fulfillment of the requirements for the degree of Doctor Scientiarum (Dr. Scient.)
References in corpus (27)
- Distilling the Knowledge in a Neural Network
- Natural Language Processing (almost) from Scratch
- Frustratingly Easy Domain Adaptation
- Theano: new features and speed improvements
- CTRL: A Conditional Transformer Language Model for Controllable Generation
- SummaRuNNer: A Recurrent Neural Network based Sequence Model for Extractive Summarization of Documents
- Automatic Detection of Fake News
- Ask the GRU: Multi-Task Learning for Deep Text Recommendations
- Clustered Multi-Task Learning: A Convex Formulation
- Learning Task Grouping and Overlap in Multi-task Learning
- BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions
- The (Un)reliability of saliency methods
- Learning Reporting Dynamics during Breaking News for Rumour Detection in Social Media
- Discourse-Aware Rumour Stance Classification in Social Media Using Sequential Classifiers
- XFake: Explainable Fake News Detector with Visualizations
- Bayesian Multitask Learning with Latent Hierarchies
- Detecting Cross-Modal Inconsistency to Defend Against Neural Fake News
- Popularity of arXiv.org within Computer Science
- Recurrent Neural Network-Based Sentence Encoder with Gated Attention for Natural Language Inference
- Adversarial Domain Adaptation for Stance Detection
- On the Importance of Delexicalization for Fact Verification
- Time-Aware Evidence Ranking for Fact-Checking
- Distantly Supervised Named Entity Recognition using Positive-Unlabeled Learning
- Universal Adversarial Perturbation for Text Classification
- Combining Sentiment Lexica with a Multi-View Variational Autoencoder
- 2kenize: Tying Subword Sequences for Chinese Script Conversion
- Neural Check-Worthiness Ranking with Weak Supervision: Finding Sentences for Fact-Checking