1 citations · 1 across the 3 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2025
A Principled Path to Fitted Distributional Evaluation
Sungee Hong, Jiayi Wang, Zhengling Qi +1
In reinforcement learning, distributional off-policy evaluation (OPE) focuses on estimating the return distribution of a target policy using offline data collected under a differen…
stat.ML2021★ 1 cited
Matrix Completion with Model-free Weighting
Jiayi Wang, Raymond K. W. Wong, Xiaojun Mao +1
In this paper, we propose a novel method for matrix completion under general non-uniform missing structures. By controlling an upper bound of a novel balancing error, we construct…