1 paper
Pranav Agarwal, Pierre de Beaucorps, Raoul de Charette
Deep reinforcement Learning for end-to-end driving is limited by the need of complex reward engineering. Sparse rewards can circumvent this challenge but suffers from long training…