Showing stat.MLShow all
3 papers · 1 filter
stat.ML2025
A Review of Causal Decision Making
Lin Ge, Hengrui Cai, Runzhe Wan +2
To make effective decisions, it is important to have a thorough understanding of the causal relationships among actions, environments, and outcomes. This review aims to surface thr…
stat.ML2025
Zero-Inflated Bandits
Haoyu Wei, Runzhe Wan, Lei Shi +1
Many real-world bandit applications are characterized by sparse rewards, which can significantly hinder learning efficiency. Leveraging problem-specific structures for careful dist…
stat.ML2024
STEEL: Singularity-aware Reinforcement Learning
Xiaohong Chen, Zhengling Qi, Runzhe Wan
Batch reinforcement learning (RL) aims at leveraging pre-collected data to find an optimal policy that maximizes the expected total rewards in a dynamic environment. The existing m…