Generative AI in Health Economics and Outcomes Research: A Taxonomy of Key Definitions and Emerging Applications, an ISPOR Working Group Report
arXiv:2410.20204 · doi:10.1016/j.jval.2025.04.2167
Abstract
Objective: This article offers a taxonomy of generative artificial intelligence (AI) for health economics and outcomes research (HEOR), explores its emerging applications, and outlines methods to enhance the accuracy and reliability of AI-generated outputs. Methods: The review defines foundational generative AI concepts and highlights current HEOR applications, including systematic literature reviews, health economic modeling, real-world evidence generation, and dossier development. Approaches such as prompt engineering (zero-shot, few-shot, chain-of-thought, persona pattern prompting), retrieval-augmented generation, model fine-tuning, and the use of domain-specific models are introduced to improve AI accuracy and reliability. Results: Generative AI shows significant potential in HEOR, enhancing efficiency, productivity, and offering novel solutions to complex challenges. Foundation models are promising in automating complex tasks, though challenges remain in scientific reliability, bias, interpretability, and workflow integration. The article discusses strategies to improve the accuracy of these AI tools. Conclusion: Generative AI could transform HEOR by increasing efficiency and accuracy across various applications. However, its full potential can only be realized by building HEOR expertise and addressing the limitations of current AI technologies. As AI evolves, ongoing research and innovation will shape its future role in the field.
36 pages, 1 figure, 2 tables
References in corpus (26)
- A Survey of Large Language Models
- Gemini: A Family of Highly Capable Multimodal Models
- Retrieval-Augmented Generation for Large Language Models: A Survey
- Publicly Available Clinical BERT Embeddings
- Reading Race: AI Recognises Patient's Racial Identity In Medical Images
- Fairness in Machine Learning: A Survey
- A Study of Generative Large Language Model for Medical Research and Healthcare
- The Rise and Potential of Large Language Model Based Agents: A Survey
- Can large language models replace humans in the systematic review process? Evaluating GPT-4's efficacy in screening and extracting data from peer-reviewed and grey literature in multiple languages
- Automated Paper Screening for Clinical Reviews Using Large Language Models
- A survey of recent methods for addressing AI fairness and bias in biomedicine
- The Prompt Report: A Systematic Survey of Prompt Engineering Techniques
- Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
- A Multi-Center Study on the Adaptability of a Shared Foundation Model for Electronic Health Records
- Advancing Multimodal Medical Capabilities of Gemini
- RA-DIT: Retrieval-Augmented Dual Instruction Tuning
- Bio-SIEVE: Exploring Instruction Tuning Large Language Models for Systematic Review Automation
- Advances and Open Challenges in Federated Foundation Models
- ELEVATE-GenAI: Reporting Guidelines for the Use of Large Language Models in Health Economics and Outcomes Research: an ISPOR Working Group on Generative AI Report
- Privacy and Security Implications of Cloud-Based AI Services : A Survey
- Advancing Real-time Pandemic Forecasting Using Large Language Models: A COVID-19 Case Study
- MAmmoTH2: Scaling Instructions from the Web
- Automatically Extracting Numerical Results from Randomized Controlled Trials with Large Language Models
- The Dark Side of Function Calling: Pathways to Jailbreaking Large Language Models
- Fine-tuning Large Language Models with Limited Data: A Survey and Practical Guide
- One Agent Too Many: User Perspectives on Approaches to Multi-agent Conversational AI