Deep Learning for Hate Speech Detection in Tweets
arXiv:1706.00188 · doi:10.1145/3041021.3054223
Abstract
Hate speech detection on Twitter is critical for applications like controversial event extraction, building AI chatterbots, content recommendation, and sentiment analysis. We define this task as being able to classify a tweet as racist, sexist or neither. The complexity of the natural language constructs makes this task very challenging. We perform extensive experiments with multiple deep learning architectures to learn semantic word embeddings to handle this complexity. Our experiments on a benchmark dataset of 16K annotated tweets show that such deep learning methods outperform state-of-the-art char/word n-gram methods by ~18 F1 points.
In Proceedings of ACM WWW'17 Companion, Perth, Western Australia, Apr 2017 (WWW'17), 2 pages
Cited by in corpus (32)
- Detecting Offensive Language in Tweets Using Deep Learning
- Cyberbullying Detection for Low-resource Languages and Dialects: Review of the State of the Art
- Stereotypical Bias Removal for Hate Speech Detection Task using Knowledge-based Generalizations
- "HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media
- Disentangling Hate in Online Memes
- An Online Multilingual Hate speech Recognition System
- Transfer Learning for Hate Speech Detection in Social Media
- Hate Speech Detection in Roman Urdu
- Multi-modal Hate Speech Detection using Machine Learning
- SoK: Content Moderation in Social Media, from Guidelines to Enforcement, and Research to Practice
- Towards countering hate speech against journalists on social media
- Denoising Multi-Source Weak Supervision for Neural Text Classification
- Hate speech detection using static BERT embeddings
- A Large-scale Dataset for Hate Speech Detection on Vietnamese Social Media Texts
- ALONE: A Dataset for Toxic Behavior among Adolescents on Twitter
- Misogynistic Tweet Detection: Modelling CNN with Small Datasets
- Deradicalizing YouTube: Characterization, Detection, and Personalization of Religiously Intolerant Arabic Videos
- Exploring multi-task multi-lingual learning of transformer models for hate speech and offensive speech identification in social media
- Evaluation of Deep Learning Models for Hostility Detection in Hindi Text
- User Identity Linkage in Social Media Using Linguistic and Social Interaction Features
- Automated Detection of Doxing on Twitter
- Graph embeddings for Abusive Language Detection
- Character-level HyperNetworks for Hate Speech Detection
- SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection
- From Universal Language Model to Downstream Task: Improving RoBERTa-Based Vietnamese Hate Speech Detection
- Hateful People or Hateful Bots? Detection and Characterization of Bots Spreading Religious Hatred in Arabic Social Media
- Deep Convolutional Neural Network Ensembles using ECOC
- Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion
- An Effective, Robust and Fairness-aware Hate Speech Detection Framework
- Hate Speech in the Political Discourse on Social Media: Disparities Across Parties, Gender, and Ethnicity
- Informed Machine Learning, Centrality, CNN, Relevant Document Detection, Repatriation of Indigenous Human Remains
- The Language of Influence: Sentiment, Emotion, and Hate Speech in State Sponsored Influence Operations