6 citations · 6 across the 2 of their papers we have counts for
3 papers · 1 filter
Directly Estimating the Variance of the λ-Return Using Temporal-Difference Methods
Craig Sherstan, Brendan Bennett, Kenny Young +4
This paper investigates estimating the variance of a temporal-difference learning agent's update target. Most reinforcement learning methods use an estimate of the value function,…
Communicative Capital for Prosthetic Agents
Patrick M. Pilarski, Richard S. Sutton, Kory W. Mathewson +3
This work presents an overarching perspective on the role that machine intelligence can play in enhancing human abilities, especially those that have been diminished due to injury…
Introspective Agents: Confidence Measures for General Value Functions
Craig Sherstan, Adam White, Marlos C. Machado +1
Agents of general intelligence deployed in real-world scenarios must adapt to ever-changing environmental conditions. While such adaptive agents may leverage engineered knowledge,…