2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2019
Hill Climbing on Value Estimates for Search-control in Dyna
Yangchen Pan, Hengshuai Yao, Amir-massoud Farahmand +1
Dyna is an architecture for model-based reinforcement learning (RL), where simulated experience from a model is used to update policies or value functions. A key component of Dyna…
cs.LG2017★ 2 cited
Effective sketching methods for value function approximation
Yangchen Pan, Erfan Sadeqi Azer, Martha White
High-dimensional representations, such as radial basis function networks or tile coding, are common choices for policy evaluation in reinforcement learning. Learning with such high…