Dynamic Measurement Scheduling for Event Forecasting using Deep RL
arXiv:1901.09699
Abstract
Imagine a patient in critical condition. What and when should be measured to forecast detrimental events, especially under the budget constraints? We answer this question by deep reinforcement learning (RL) that jointly minimizes the measurement cost and maximizes predictive gain, by scheduling strategically-timed measurements. We learn our policy to be dynamically dependent on the patient's health history. To scale our framework to exponentially large action space, we distribute our reward in a sequential setting that makes the learning easier. In our simulation, our policy outperforms heuristic-based scheduling with higher predictive gain and lower cost. In a real-world ICU mortality prediction task (MIMIC3), our policies reduce the total number of measurements by or improve predictive gain by a factor of as compared to physicians, under the off-policy policy evaluation.
ICML 2019
References in corpus (5)
- Continuous State-Space Models for Optimal Sepsis Treatment - a Deep Reinforcement Learning Approach
- An Improved Multi-Output Gaussian Process RNN with Real-Time Validation for Early Sepsis Detection
- Representation and Reinforcement Learning for Personalized Glycemic Control in Septic Patients
- Medical Diagnosis From Laboratory Tests by Combining Generative and Discriminative Learning
- Why Pay More When You Can Pay Less: A Joint Learning Framework for Active Feature Acquisition and Classification