1 paper · 1 filter
Zeyu Bian, Chengchun Shi, Zhengling Qi +1
This work aims to study off-policy evaluation (OPE) under scenarios where two key reinforcement learning (RL) assumptions -- temporal stationarity and individual homogeneity are bo…