1 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2021★ 1 cited
An Entropy Regularization Free Mechanism for Policy-based Reinforcement Learning
Changnan Xiao, Haosen Shi, Jiajun Fan +1
Policy-based reinforcement learning methods suffer from the policy collapse problem. We find valued-based reinforcement learning methods with ε-greedy mechanism are capable of enjo…
cs.LG2020★ 1 cited
Critic PI2: Master Continuous Planning via Policy Improvement with Path Integrals and Deep Actor-Critic Reinforcement Learning
Jiajun Fan, He Ba, Xian Guo +1
Constructing agents with planning capabilities has long been one of the main challenges in the pursuit of artificial intelligence. Tree-based planning methods from AlphaGo to Muzer…