230 citations · 244 across the 10 of their papers we have counts for
1 paper · 1 filter
Aviv Tamar, Panos Toulis, Shie Mannor +1
In reinforcement learning, the TD(λ) algorithm is a fundamental policy evaluation method with an efficient online implementation that is suitable for large-scale problems. One pr…