7 citations · 7 across the 2 of their papers we have counts for
1 paper · 1 filter
Aviv Tamar, Panos Toulis, Shie Mannor +1
In reinforcement learning, the TD(λ) algorithm is a fundamental policy evaluation method with an efficient online implementation that is suitable for large-scale problems. One pr…