2 citations · 2 across the 1 of their papers we have counts for
2 papers
cs.LG2019
FiDi-RL: Incorporating Deep Reinforcement Learning with Finite-Difference Policy Search for Efficient Learning of Continuous Control
Longxiang Shi, Shijian Li, Longbing Cao +3
In recent years significant progress has been made in dealing with challenging problems using reinforcement learning.Despite its great success, reinforcement learning still faces c…
cs.LG2019★ 2 cited
TBQ(): Improving Efficiency of Trace Utilization for Off-Policy Reinforcement Learning
Longxiang Shi, Shijian Li, Longbing Cao +2
Off-policy reinforcement learning with eligibility traces is challenging because of the discrepancy between target policy and behavior policy. One common approach is to measure the…