1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Qian Chen, Junqiao Zhao, Hongtu Zhou +4
Long-horizon, sparse-reward tasks pose a fundamental challenge for reinforcement learning, since single-step TD learning suffers from bootstrapping error accumulation across succes…