3 citations · 7 across the 4 of their papers we have counts for
4 papers
ε-Neural Thompson Sampling of Deep Brain Stimulation for Parkinson Disease Treatment
Hao-Lun Hsu, Qitong Gao, Miroslav Pajic
Deep Brain Stimulation (DBS) stands as an effective intervention for alleviating the motor symptoms of Parkinson's disease (PD). Traditional commercial DBS devices are only able to…
Off-Policy Evaluation for Human Feedback
Qitong Gao, Ge Gao, Juncheng Dong +3
Off-policy evaluation (OPE) is important for closing the gap between offline training and evaluation of reinforcement learning (RL), by estimating performance and/or rank of target…
Robust Reinforcement Learning through Efficient Adversarial Herding
Juncheng Dong, Hao-Lun Hsu, Qitong Gao +2
Although reinforcement learning (RL) is considered the gold standard for policy design, it may not always provide a robust solution in various scenarios. This can result in severe…
Variational Latent Branching Model for Off-Policy Evaluation
Qitong Gao, Ge Gao, Min Chi +1
Model-based methods have recently shown great potential for off-policy evaluation (OPE); offline trajectories induced by behavioral policies are fitted to transitions of Markov dec…