1 paper
Shani Gamrian, Yoav Goldberg
Despite the remarkable success of Deep RL in learning control policies from raw pixels, the resulting models do not generalize. We demonstrate that a trained agent fails completely…