1 paper
Yu-Jui Huang, Zhenhua Wang, Zhou Zhou
For a general entropy-regularized stochastic control problem on an infinite horizon, we prove that a policy iteration algorithm (PIA) converges to an optimal relaxed control. Contr…