2 papers
stat.ML2026
Distributional Off-Policy Evaluation with Deep Quantile Process Regression
Qi Kuang, Chao Wang, Yuling Jiao +1
This paper investigates the off-policy evaluation (OPE) problem from a distributional perspective. Rather than focusing solely on the expectation of the total return, as in most ex…
cs.LG2024
Two-way Deconfounder for Off-policy Evaluation in Causal Reinforcement Learning
Shuguang Yu, Shuxing Fang, Ruixin Peng +3
This paper studies off-policy evaluation (OPE) in the presence of unmeasured confounders. Inspired by the two-way fixed effects regression model widely used in the panel data liter…