5 citations · 7 across the 5 of their papers we have counts for
1 paper · 1 filter
Bowen He, Sreehari Rammohan, Jessica Forde +1
In this work, we study two self-play training schemes, Chainer and Pool, and show they lead to improved agent performance in Atari Pong compared to a standard DQN agent -- trained…