A Domain-adaptive Pre-training Approach for Language Bias Detection in News
arXiv:2205.10773 · doi:10.1145/3529372.3530932
Abstract
Media bias is a multi-faceted construct influencing individual behavior and collective decision-making. Slanted news reporting is the result of one-sided and polarized writing which can occur in various forms. In this work, we focus on an important form of media bias, i.e. bias by word choice. Detecting biased word choices is a challenging task due to its linguistic complexity and the lack of representative gold-standard corpora. We present DA-RoBERTa, a new state-of-the-art transformer-based model adapted to the media bias domain which identifies sentence-level bias with an F1 score of 0.814. In addition, we also train, DA-BERT and DA-BART, two more transformer models adapted to the bias domain. Our proposed domain-adapted models outperform prior bias detection approaches on the same data.
References in corpus (4)
- Multi-task learning for natural language processing in the 2020s: where are we going?
- Enabling News Consumers to View and Understand Biased News Coverage: A Study on the Perception and Visualization of Media Bias
- Exploiting Transformer-based Multitask Learning for the Detection of Media Bias in News Articles
- Identification of Biased Terms in News Articles by Comparison of Outlet-specific Word Embeddings