4 citations · 10 across the 3 of their papers we have counts for
3 papers
cs.LG2020★ 3 cited
Proximal Policy Optimization via Enhanced Exploration Efficiency
Junwei Zhang, Zhenghao Zhang, Shuai Han +1
Proximal policy optimization (PPO) algorithm is a deep reinforcement learning algorithm with outstanding performance, especially in continuous control tasks. But the performance of…
cs.LG2020★ 4 cited
NROWAN-DQN: A Stable Noisy Network with Noise Reduction and Online Weight Adjustment for Exploration
Shuai Han, Wenbo Zhou, Jing Liu +1
Deep reinforcement learning has been applied more and more widely nowadays, especially in various complex control tasks. Effective exploration for noisy networks is one of the most…
cs.LG2019★ 3 cited
Recruitment-imitation Mechanism for Evolutionary Reinforcement Learning
Shuai Lü, Shuai Han, Wenbo Zhou +1
Reinforcement learning, evolutionary algorithms and imitation learning are three principal methods to deal with continuous control tasks. Reinforcement learning is sample efficient…