1 paper
Yingrong Wang, Anpeng Wu, Haoxuan Li +5
This paper focuses on developing Pareto-optimal estimation and policy learning to identify the most effective treatment that maximizes the total reward from both short-term and lon…