1 paper
Christian Bender, Nguyen Tran Thuan
Motivated by the trade-off between exploitation and exploration in reinforcement learning, we study a continuous-time entropy-regularized mean variance portfolio selection problem…