3 papers
stat.ML2024
Causal Deepsets for Off-policy Evaluation under Spatial or Spatio-temporal Interferences
Runpeng Dai, Jianing Wang, Fan Zhou +4
Off-policy evaluation (OPE) is widely applied in sectors such as pharmaceuticals and e-commerce to evaluate the efficacy of novel products or policies from offline datasets. This p…
stat.ME2023
Optimal Treatment Allocation for Efficient Policy Evaluation in Sequential Decision Making
Ting Li, Chengchun Shi, Jianing Wang +2
A/B testing is critical for modern technological companies to evaluate the effectiveness of newly developed products against standard baselines. This paper studies optimal designs…
stat.ML2023
Value Enhancement of Reinforcement Learning via Efficient and Robust Trust Region Optimization
Chengchun Shi, Zhengling Qi, Jianing Wang +1
Reinforcement learning (RL) is a powerful machine learning technique that enables an intelligent agent to learn an optimal policy that maximizes the cumulative rewards in sequentia…