15 citations · 37 across the 3 of their papers we have counts for
3 papers
cs.AI2016★ 14 cited
Accelerated Gradient Temporal Difference Learning
Yangchen Pan, Adam White, Martha White
The family of temporal difference (TD) methods span a spectrum from computationally frugal linear methods like TD(λ) to data efficient least squares methods. Least square methods m…
cs.AI2016★ 8 cited
A Greedy Approach to Adapting the Trace Parameter for Temporal Difference Learning
Martha White, Adam White
One of the main obstacles to broad application of reinforcement learning methods is the parameter sensitivity of our core learning algorithms. In many large-scale applications, onl…
cs.LG2016★ 15 cited
Investigating practical linear temporal difference learning
Adam White, Martha White
Off-policy reinforcement learning has many applications including: learning from demonstration, learning multiple goal seeking policies in parallel, and representing predictive kno…