activity
20192022
most citedOffline Reinforcement Learning with Instrumental Variables in Confounded Markov Decision Processes

6 citations · 14 across the 6 of their papers we have counts for

collaborators

7 papers

cs.LG20223 cited

RISE: Robust Individualized Decision Learning with Sensitive Variables

Xiaoqing Tan, Zhengling Qi, Christopher W. Seymour +1

This paper introduces RISE, a robust individualized decision learning framework with sensitive variables, where sensitive variables are collectible data and important to the interv…

stat.ML20221 cited

Off-Policy Evaluation for Episodic Partially Observable Markov Decision Processes under Non-Parametric Models

Rui Miao, Zhengling Qi, Xiaoke Zhang

We study the problem of off-policy evaluation (OPE) for episodic Partially Observable Markov Decision Processes (POMDPs) with continuous states. Motivated by the recently proposed…

cs.LG20226 cited

Offline Reinforcement Learning with Instrumental Variables in Confounded Markov Decision Processes

Zuyue Fu, Zhengling Qi, Zhaoran Wang +3

We study the offline reinforcement learning (RL) in the face of unmeasured confounders. Due to the lack of online interaction with the environment, offline RL is facing the followi…

stat.ML20211 cited

Rejoinder: Learning Optimal Distributionally Robust Individualized Treatment Rules

Weibin Mo, Zhengling Qi, Yufeng Liu

We thank the opportunity offered by editors for this discussion and the discussants for their insightful comments and thoughtful contributions. We also want to congratulate Kallus…

stat.ML2020

Learning Optimal Distributionally Robust Individualized Treatment Rules

Weibin Mo, Zhengling Qi, Yufeng Liu

Recent development in the data-driven decision science has seen great advances in individualized decision making. Given data with individual covariates, treatment assignments and o…

math.ST20193 cited

Statistical Analysis of Stationary Solutions of Coupled Nonconvex Nonsmooth Empirical Risk Minimization

Zhengling Qi, Ying Cui, Yufeng Liu +1

This paper has two main goals: (a) establish several statistical properties---consistency, asymptotic distributions, and convergence rates---of stationary solutions and values of a…