14 citations · 14 across the 1 of their papers we have counts for
1 paper
Beining Han, Zhizhou Ren, Zuofan Wu +2
We study deep reinforcement learning (RL) algorithms with delayed rewards. In many real-world tasks, instant rewards are often not readily accessible or even defined immediately af…