668 citations · 875 across the 13 of their papers we have counts for
3 papers · 1 filter
Imitation by Predicting Observations
Andrew Jaegle, Yury Sulsky, Arun Ahuja +3
Imitation learning enables agents to reuse and adapt the hard-won expertise of others, offering a solution to several key challenges in learning behavior. Although it is easy to ob…
Synthetic Returns for Long-Term Credit Assignment
David Raposo, Sam Ritter, Adam Santoro +5
Since the earliest days of reinforcement learning, the workhorse method for assigning credit to actions over time has been temporal-difference (TD) learning, which propagates credi…
Imitating Interactive Intelligence
Josh Abramson, Arun Ahuja, Iain Barr +26
A common vision from science fiction is that robots will one day inhabit our physical spaces, sense the world as we do, assist our physical labours, and communicate with us through…