3 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.LG2021★ 1 cited
ADER:Adapting between Exploration and Robustness for Actor-Critic Methods
Bo Zhou, Kejiao Li, Hongsheng Zeng +2
Combining off-policy reinforcement learning methods with function approximators such as neural networks has been found to lead to overestimation of the value function and sub-optim…
cs.AI2021★ 3 cited
Action Set Based Policy Optimization for Safe Power Grid Management
Bo Zhou, Hongsheng Zeng, Yuecheng Liu +3
Maintaining the stability of the modern power grid is becoming increasingly difficult due to fluctuating power consumption, unstable power supply coming from renewable energies, an…