8 citations · 8 across the 2 of their papers we have counts for
1 paper · 1 filter
Lingheng Meng, Rob Gorbet, Dana Kulić
Multi-step (also called n-step) methods in reinforcement learning (RL) have been shown to be more efficient than the 1-step method due to faster propagation of the reward signal, b…