Survey and Experiments on Mental Disorder Detection via Social Media: From Large Language Models and RAG to Agents
arXiv:2504.02800 · doi:10.1109/ICDEW67478.2025.00027
Abstract
Mental disorders represent a critical global health challenge, and social media is increasingly viewed as a vital resource for real-time digital phenotyping and intervention. To leverage this data, large language models (LLMs) have been introduced, offering stronger semantic understanding and reasoning than traditional deep learning, thereby enhancing the explainability of detection results. Despite the growing prominence of LLMs in this field, there is a scarcity of scholarly works that systematically synthesize how advanced enhancement techniques, specifically Retrieval-Augmented Generation (RAG) and Agentic systems, can be utilized to address these reliability and reasoning limitations. Here, we systematically survey the evolving landscape of LLM-based methods for social media mental disorder analysis, spanning standard pre-trained language models, RAG to mitigate hallucinations and contextual gaps, and agentic systems for autonomous reasoning and multi-step intervention. We organize existing work by technical paradigm and clinical target, extending beyond common internalizing disorders to include psychotic disorders and externalizing behaviors. Additionally, the paper comprehensively evaluates the performance of LLMs, including the impact of RAG, across various tasks. This work establishes a unified benchmark for the field, paving the way for the development of trustworthy, autonomous AI systems that can deliver precise and explainable mental health support.
20 pages, 10 figures. This is an extension of ICDEW 2025
References in corpus (11)
- LLaMA: Open and Efficient Foundation Language Models
- DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
- Socially Aware Synthetic Data Generation for Suicidal Ideation Detection Using Large Language Models
- Evaluation of ChatGPT for NLP-based Mental Health Applications
- DeepSeek-VL: Towards Real-World Vision-Language Understanding
- A Multitask Deep Learning Approach for User Depression Detection on Sina Weibo
- Read, Diagnose and Chat: Towards Explainable and Interactive LLMs-Augmented Depression Detection in Social Media
- A Survey on Data Synthesis and Augmentation for Large Language Models
- Applying and Evaluating Large Language Models in Mental Health Care: A Scoping Review of Human-Assessed Generative Tasks
- Surveying the Effects of Quality, Diversity, and Complexity in Synthetic Data From Large Language Models
- A Survey on Large Language Model Acceleration based on KV Cache Management