2 papers
cs.LG2025
Generalized Fitted Q-Iteration with Clustered Data
Liyuan Hu, Jitao Wang, Zhenke Wu +1
This paper focuses on reinforcement learning (RL) with clustered data, which is commonly encountered in healthcare applications. We propose a generalized fitted Q-iteration (FQI) a…
stat.ML2025
Off-policy Evaluation with Deeply-abstracted States
Meiling Hao, Pingfan Su, Liyuan Hu +3
Off-policy evaluation (OPE) is crucial for assessing a target policy's impact offline before its deployment. However, achieving accurate OPE in large state spaces remains challengi…