1 citations · 1 across the 1 of their papers we have counts for
1 paper
Tian Tian, Kenny Young, Richard S. Sutton
Value iteration (VI) is a foundational dynamic programming method, important for learning and planning in optimal control and reinforcement learning. VI proceeds in batches, where…