153 citations · 589 across the 46 of their papers we have counts for
Showing 2018 · cs.LGShow all
3 papers · 2 filters
cs.LG2018
Learning Curriculum Policies for Reinforcement Learning
Sanmit Narvekar, Peter Stone
Curriculum learning in reinforcement learning is a training methodology that seeks to speed up learning of a difficult target task, by first training on a series of simpler tasks a…
cs.LG2018
Generative Adversarial Imitation from Observation
Faraz Torabi, Garrett Warnell, Peter Stone
Imitation from observation (IfO) is the problem of learning directly from state-only demonstrations without having access to the demonstrator's actions. The lack of action informat…
cs.LG2018
Importance Sampling Policy Evaluation with an Estimated Behavior Policy
Josiah P. Hanna, Scott Niekum, Peter Stone
We consider the problem of off-policy evaluation in Markov decision processes. Off-policy evaluation is the task of evaluating the expected return of one policy with data generated…