1 paper
Yuhao Ding, Junzi Zhang, Hyunin Lee +1
Entropy regularization is an efficient technique for encouraging exploration and preventing a premature convergence of (vanilla) policy gradient methods in reinforcement learning (…