1 paper
Bogdan Mazoure, Walter Talbott, Miguel Angel Bautista +3
A fairly reliable trend in deep reinforcement learning is that the performance scales with the number of parameters, provided a complimentary scaling in amount of training data. As…