233 citations · 441 across the 25 of their papers we have counts for
Showing 2018 · cs.LGShow all
2 papers · 2 filters
cs.LG2018
One-Shot High-Fidelity Imitation: Training Large-Scale Deep Nets with RL
Tom Le Paine, Sergio Gómez Colmenarejo, Ziyu Wang +8
Humans are experts at high-fidelity imitation -- closely mimicking a demonstration, often in one attempt. Humans use this ability to quickly solve a task instance, and to bootstrap…
cs.LG2018
Playing hard exploration games by watching YouTube
Yusuf Aytar, Tobias Pfaff, David Budden +3
Deep reinforcement learning methods traditionally struggle with tasks where environment rewards are particularly sparse. One successful method of guiding exploration in these domai…