Patient Outcome and Zero-shot Diagnosis Prediction with Hypernetwork-guided Multitask Learning
arXiv:2109.03062
Abstract
Multitask deep learning has been applied to patient outcome prediction from text, taking clinical notes as input and training deep neural networks with a joint loss function of multiple tasks. However, the joint training scheme of multitask learning suffers from inter-task interference, and diagnosis prediction among the multiple tasks has the generalizability issue due to rare diseases or unseen diagnoses. To solve these challenges, we propose a hypernetwork-based approach that generates task-conditioned parameters and coefficients of multitask prediction heads to learn task-specific prediction and balance the multitask learning. We also incorporate semantic task information to improves the generalizability of our task-conditioned multitask model. Experiments on early and discharge notes extracted from the real-world MIMIC database show our method can achieve better performance on multitask patient outcome prediction than strong baselines in most cases. Besides, our method can effectively handle the scenario with limited information and improve zero-shot prediction on unseen diagnosis categories.
EACL 2023
References in corpus (5)
- Adam: A Method for Stochastic Optimization
- Deep EHR: A Survey of Recent Advances in Deep Learning Techniques for Electronic Health Record (EHR) Analysis
- ClinicalBERT: Modeling Clinical Notes and Predicting Hospital Readmission
- Phenotyping of Clinical Notes with Improved Document Classification Models Using Contextualized Neural Language Models
- Multitask Recalibrated Aggregation Network for Medical Code Prediction