64 citations · 139 across the 5 of their papers we have counts for
3 papers · 1 filter
Task-Relevant Adversarial Imitation Learning
Konrad Zolna, Scott Reed, Alexander Novikov +6
We show that a critical vulnerability in adversarial imitation is the tendency of discriminator networks to learn spurious associations between visual features and expert labels. W…
Scaling data-driven robotics with reward sketching and batch reinforcement learning
Serkan Cabi, Sergio Gómez Colmenarejo, Alexander Novikov +13
We present a framework for data-driven robotics that makes use of a large dataset of recorded robot experience and scales to several tasks using learned reward functions. We show h…
Learning Compositional Neural Programs with Recursive Tree Search and Planning
Thomas Pierrot, Guillaume Ligner, Scott Reed +6
We propose a novel reinforcement learning algorithm, AlphaNPI, that incorporates the strengths of Neural Programmer-Interpreters (NPI) and AlphaZero. NPI contributes structural bia…