Large Language Models for Healthcare Text Classification: A Systematic Review
arXiv:2503.01159 · doi:10.2196/79202
Abstract
Large Language Models (LLMs) have fundamentally transformed approaches to Natural Language Processing (NLP) tasks across diverse domains. In healthcare, accurate and cost-efficient text classification is crucial, whether for clinical notes analysis, diagnosis coding, or any other task, and LLMs present promising potential. Text classification has always faced multiple challenges, including manual annotation for training, handling imbalanced data, and developing scalable approaches. With healthcare, additional challenges are added, particularly the critical need to preserve patients' data privacy and the complexity of the medical terminology. Numerous studies have been conducted to leverage LLMs for automated healthcare text classification and contrast the results with existing machine learning-based methods where embedding, annotation, and training are traditionally required. Existing systematic reviews about LLMs either do not specialize in text classification or do not focus on the healthcare domain. This research synthesizes and critically evaluates the current evidence found in the literature regarding the use of LLMs for text classification in a healthcare setting. Major databases (e.g., Google Scholar, Scopus, PubMed, Science Direct) and other resources were queried, which focused on the papers published between 2018 and 2024 within the framework of PRISMA guidelines, which resulted in 65 eligible research articles. These were categorized by text classification type (e.g., binary classification, multi-label classification), application (e.g., clinical decision support, public health and opinion analysis), methodology, type of healthcare text, and metrics used for evaluation and validation. This review reveals the existing gaps in the literature and suggests future research lines that can be investigated and explored.
References in corpus (11)
- Machine Learning in Automated Text Categorization
- BioBERT: a pre-trained biomedical language representation model for biomedical text mining
- Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data
- Automated Paper Screening for Clinical Reviews Using Large Language Models
- A Comparative Study of Pretrained Language Models for Long Clinical Text
- Taiyi: A Bilingual Fine-Tuned Large Language Model for Diverse Biomedical Tasks
- Evaluation of ChatGPT Family of Models for Biomedical Reasoning and Classification
- Two Directions for Clinical Data Generation with Large Language Models: Data-to-Label and Label-to-Data
- Predicting Clinical Diagnosis from Patients Electronic Health Records Using BERT-based Neural Networks
- DALLMi: Domain Adaption for LLM-based Multi-label Classifier
- Predicting COVID-19 Patient Shielding: A Comprehensive Study