2 citations · 2 across the 1 of their papers we have counts for
1 paper
Anna Winnicki, R. Srikant
A common technique in reinforcement learning is to evaluate the value function from Monte Carlo simulations of a given policy, and use the estimated value function to obtain a new…