2 citations · 2 across the 1 of their papers we have counts for
1 paper
Haoya Li, Samarth Gupta, Hsiangfu Yu +2
Policy gradient algorithms have been widely applied to Markov decision processes and reinforcement learning problems in recent years. Regularization with various entropy functions…