39 citations · 77 across the 10 of their papers we have counts for
1 paper · 1 filter
Christoph Dann, Chen-Yu Wei, Julian Zimmert
We study reinforcement learning in stochastic path (SP) problems. The goal in these problems is to maximize the expected sum of rewards until the agent reaches a terminal state. We…