1 citations · 1 across the 1 of their papers we have counts for
1 paper
Changnan Xiao, Haosen Shi, Jiajun Fan +1
Policy-based reinforcement learning methods suffer from the policy collapse problem. We find valued-based reinforcement learning methods with ε-greedy mechanism are capable of enjo…