53 citations · 74 across the 6 of their papers we have counts for
Showing 2019Show all
2 papers · 1 filter
cs.LG2019
Evaluating task-agnostic exploration for fixed-batch learning of arbitrary future tasks
Vibhavari Dasagi, Robert Lee, Jake Bruce +1
Deep reinforcement learning has been shown to solve challenging tasks where large amounts of training experience is available, usually obtained online while learning the task. Robo…
cs.LG2019★ 13 cited
Ctrl-Z: Recovering from Instability in Reinforcement Learning
Vibhavari Dasagi, Jake Bruce, Thierry Peynot +1
When learning behavior, training data is often generated by the learner itself; this can result in unstable training dynamics, and this problem has particularly important applicati…