3 papers
math.ST2026
Semiparametric Off-Policy Inference for Optimal Policy Values under Possible Non-Uniqueness
Haoyu Wei
Off-policy evaluation (OPE) constructs confidence intervals for the value of a target policy using data generated under a different behavior policy. Most existing inference methods…
econ.EM2026
Difference-in-differences with a mediator
Yuhao Deng, Haoyu Wei, Zhongzhe Ouyang
Causal mediation analysis is a powerful tool for disentangling the total effect of a treatment into its direct effect on the outcome and its indirect effect mediated through an int…
stat.ML2025
Zero-Inflated Bandits
Haoyu Wei, Runzhe Wan, Lei Shi +1
Many real-world bandit applications are characterized by sparse rewards, which can significantly hinder learning efficiency. Leveraging problem-specific structures for careful dist…