1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Henrik Müller, Daniel Kudenko
Potential-based reward shaping is commonly used to incorporate prior knowledge of how to solve the task into reinforcement learning because it can formally guarantee policy invaria…