1 paper
Aditya A. Ramesh, Kenny Young, Louis Kirsch +1
Temporal credit assignment in reinforcement learning is challenging due to delayed and stochastic outcomes. Monte Carlo targets can bridge long delays between action and consequenc…