1 paper
Yihong Guo, Hao Liu, Yisong Yue +1
We introduce a distributionally robust approach that enhances the reliability of offline policy evaluation in contextual bandits under general covariate shifts. Our method aims to…