1 citations · 1 across the 2 of their papers we have counts for
3 papers
stat.ML2023
Distributional Shift-Aware Off-Policy Interval Estimation: A Unified Error Quantification Framework
Wenzhuo Zhou, Yuhan Li, Ruoqing Zhu +1
We study high-confidence off-policy evaluation in the context of infinite-horizon Markov decision processes, where the objective is to establish a confidence interval (CI) for the…
stat.ME2023★ 1 cited
Policy Learning for Individualized Treatment Regimes on Infinite Time Horizon
Wenzhuo Zhou, Yuhan Li, Ruoqing Zhu
With the recent advancements of technology in facilitating real-time monitoring and data collection, "just-in-time" interventions can be delivered via mobile devices to achieve bot…
stat.ML2023
Quasi-optimal Reinforcement Learning with Continuous Actions
Yuhan Li, Wenzhuo Zhou, Ruoqing Zhu
Many real-world applications of reinforcement learning (RL) require making decisions in continuous action environments. In particular, determining the optimal dose level plays a vi…