1 paper
Anthony Liang, Jesse Thomason, Erdem Bıyık
Training robots to perform complex control tasks from high-dimensional pixel input using reinforcement learning (RL) is sample-inefficient, because image observations are comprised…