2 papers
stat.ML2025
Missing Data Multiple Imputation for Tabular Q-Learning in Online RL
Kyla Chasalow, Skyler Wu, Susan Murphy
Missing data in online reinforcement learning (RL) poses challenges compared to missing data in standard tabular data or in offline policy learning. The need to impute and act at e…
cs.LG2024
BOTS: Batch Bayesian Optimization of Extended Thompson Sampling for Severely Episode-Limited RL Settings
Karine Karine, Susan A. Murphy, Benjamin M. Marlin
In settings where the application of reinforcement learning (RL) requires running real-world trials, including the optimization of adaptive health interventions, the number of epis…