24 citations · 29 across the 4 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2020
Riemannian Proximal Policy Optimization
Shijun Wang, Baocheng Zhu, Chen Li +4
In this paper, We propose a general Riemannian proximal optimization algorithm with guaranteed convergence to solve Markov decision process (MDP) problems. To model policy function…
cs.LG2018
Reinforcement Learning for Uplift Modeling
Chenchen Li, Xiang Yan, Xiaotie Deng +6
Uplift modeling aims to directly model the incremental impact of a treatment on an individual response. In this work, we address the problem from a new angle and reformulate it as…