2 citations · 2 across the 6 of their papers we have counts for
1 paper · 1 filter
Hans Harder, Sebastian Peitz
The value function plays a crucial role as a measure for the cumulative future reward an agent receives in both reinforcement learning and optimal control. It is therefore of inter…