6 citations · 6 across the 4 of their papers we have counts for
3 papers · 1 filter
Work in Progress: Temporally Extended Auxiliary Tasks
Craig Sherstan, Bilal Kartal, Pablo Hernandez-Leal +1
Predictive auxiliary tasks have been shown to improve performance in numerous reinforcement learning works, however, this effect is still not well understood. The primary purpose o…
Gamma-Nets: Generalizing Value Estimation over Timescale
Craig Sherstan, Shibhansh Dohare, James MacGlashan +2
We present -nets, a method for generalizing value function estimation over timescale. By using the timescale as one of the estimator's inputs we can estimate value for arbitrary…
Accelerating Learning in Constructive Predictive Frameworks with the Successor Representation
Craig Sherstan, Marlos C. Machado, Patrick M. Pilarski
Here we propose using the successor representation (SR) to accelerate learning in a constructive knowledge system based on general value functions (GVFs). In real-world settings li…