Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection
arXiv:2309.12247 · doi:10.1609/aaai.v38i20.30214
Abstract
Detecting fake news requires both a delicate sense of diverse clues and a profound understanding of the real-world background, which remains challenging for detectors based on small language models (SLMs) due to their knowledge and capability limitations. Recent advances in large language models (LLMs) have shown remarkable performance in various tasks, but whether and how LLMs could help with fake news detection remains underexplored. In this paper, we investigate the potential of LLMs in fake news detection. First, we conduct an empirical study and find that a sophisticated LLM such as GPT 3.5 could generally expose fake news and provide desirable multi-perspective rationales but still underperforms the basic SLM, fine-tuned BERT. Our subsequent analysis attributes such a gap to the LLM's inability to select and integrate rationales properly to conclude. Based on these findings, we propose that current LLMs may not substitute fine-tuned SLMs in fake news detection but can be a good advisor for SLMs by providing multi-perspective instructive rationales. To instantiate this proposal, we design an adaptive rationale guidance network for fake news detection (ARG), in which SLMs selectively acquire insights on news analysis from the LLMs' rationales. We further derive a rationale-free version of ARG by distillation, namely ARG-D, which services cost-sensitive scenarios without querying LLMs. Experiments on two real-world datasets demonstrate that ARG and ARG-D outperform three types of baseline methods, including SLM-based, LLM-based, and combinations of small and large language models.
16 pages, 5 figures, and 9 tables. To appear at AAAI 2024
References in corpus (24)
- Distilling the Knowledge in a Neural Network
- Survey of Hallucination in Natural Language Generation
- Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
- LLaMA: Open and Efficient Foundation Language Models
- A Survey of Large Language Models
- Large Language Models are Zero-Shot Reasoners
- Emergent Abilities of Large Language Models
- ChatGPT: Jack of all trades, master of none
- Fake News Detection on Social Media: A Data Mining Perspective
- Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models
- MDFEND: Multi-domain Fake News Detection
- Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT
- Large Language Model Is Not a Good Few-shot Information Extractor, but a Good Reranker for Hard Samples!
- Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
- Domain Adaptive Fake News Detection via Reinforcement Learning
- Zoom Out and Observe: News Environment Perception for Fake News Detection
- Generalizing to the Future: Mitigating Entity Bias in Fake News Detection
- Integrating Pattern- and Fact-based Fake News Detection via Model Preference Learning
- Network-based Fake News Detection: A Pattern-driven Approach
- Small Models are Valuable Plug-ins for Large Language Models
- News Verifiers Showdown: A Comparative Performance Evaluation of ChatGPT 3.5, ChatGPT 4.0, Bing AI, and Bard in News Fact-Checking
- Towards Reliable Misinformation Mitigation: Generalization, Uncertainty, and GPT-4
- Learn over Past, Evolve for Future: Forecasting Temporal Trends for Fake News Detection
- It's about Time: Rethinking Evaluation on Rumor Detection Benchmarks using Chronological Splits
Cited by in corpus (8)
- Emotion Detection for Misinformation: A Review
- Hallucination to Truth: A Review of Fact-Checking and Factuality Evaluation in Large Language Models
- A Macro- and Micro-Hierarchical Transfer Learning Framework for Cross-Domain Fake News Detection
- LLM-Generated Fake News Induces Truth Decay in News Ecosystem: A Case Study on Neural News Recommendation
- QuestGen: Effectiveness of Question Generation Methods for Fact-Checking Applications
- Exploring news intent and its application: A theory-driven approach
- Forecasting the Buzz: Enriching Hashtag Popularity Prediction with LLM Reasoning
- Enhancing Debunking Effectiveness through LLM-based Personality Adaptation