1 paper
Shijun Wang, Baocheng Zhu, Chen Li +4
In this paper, We propose a general Riemannian proximal optimization algorithm with guaranteed convergence to solve Markov decision process (MDP) problems. To model policy function…