13 citations · 34 across the 8 of their papers we have counts for
1 paper · 1 filter
Baicen Xiao, Bhaskar Ramasubramanian, Andrew Clark +3
This paper augments the reward received by a reinforcement learning agent with potential functions in order to help the agent learn (possibly stochastic) optimal policies. We show…