3 citations · 4 across the 4 of their papers we have counts for
1 paper · 1 filter
Tingting Zhao, Hirotaka Hachiya, Voot Tangkaratt +2
The policy gradient approach is a flexible and powerful reinforcement learning method particularly for problems with continuous actions such as robot control. A common challenge in…