1 paper
Daphne Cornelisse, Julian Hunt, Zixu Zhang +4
Self-play reinforcement learning has recently emerged as a way to train driving policies without any human data. It uses cheap, large-scale simulations to substitute expensive, lar…