34 citations · 66 across the 4 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2020★ 14 cited
Policy Evaluation Networks
Jean Harb, Tom Schaul, Doina Precup +1
Many reinforcement learning algorithms use value functions to guide the search for better policies. These methods estimate the value of a single policy while generalizing across ma…
cs.LG2017★ 34 cited
Learnings Options End-to-End for Continuous Action Tasks
Martin Klissarov, Pierre-Luc Bacon, Jean Harb +1
We present new results on learning temporally extended actions for continuoustasks, using the options framework (Suttonet al.[1999b], Precup [2000]). In orderto achieve this goal w…