1 paper · 1 filter
Touqir Sajed, Wesley Chung, Martha White
Estimating the value function for a fixed policy is a fundamental problem in reinforcement learning. Policy evaluation algorithms---to estimate value functions---continue to be dev…