4 citations · 11 across the 4 of their papers we have counts for
4 papers
Proximal Policy Optimization via Enhanced Exploration Efficiency
Junwei Zhang, Zhenghao Zhang, Shuai Han +1
Proximal policy optimization (PPO) algorithm is a deep reinforcement learning algorithm with outstanding performance, especially in continuous control tasks. But the performance of…
Regularly Updated Deterministic Policy Gradient Algorithm
Shuai Han, Wenbo Zhou, Shuai Lü +1
Deep Deterministic Policy Gradient (DDPG) algorithm is one of the most well-known reinforcement learning methods. However, this method is inefficient and unstable in practical appl…
NROWAN-DQN: A Stable Noisy Network with Noise Reduction and Online Weight Adjustment for Exploration
Shuai Han, Wenbo Zhou, Jing Liu +1
Deep reinforcement learning has been applied more and more widely nowadays, especially in various complex control tasks. Effective exploration for noisy networks is one of the most…
Recruitment-imitation Mechanism for Evolutionary Reinforcement Learning
Shuai Lü, Shuai Han, Wenbo Zhou +1
Reinforcement learning, evolutionary algorithms and imitation learning are three principal methods to deal with continuous control tasks. Reinforcement learning is sample efficient…