12 citations · 24 across the 18 of their papers we have counts for
1 paper · 2 filters
Qian Chen, Junqiao Zhao, Hongtu Zhou +4
Long-horizon, sparse-reward tasks pose a fundamental challenge for reinforcement learning, since single-step TD learning suffers from bootstrapping error accumulation across succes…