1 paper · 1 filter
Hung Vinh Tran, Zhenhua Wang, Yuming Paul Zhang
We study the policy iteration algorithm (PIA) for entropy-regularized stochastic control problems on an infinite time horizon with a large discount rate, focusing on two main scena…