66 citations · 67 across the 4 of their papers we have counts for
1 paper · 1 filter
Chen Tessler, Shie Mannor
In reinforcement learning, the discount factor γ controls the agent's effective planning horizon. Traditionally, this parameter was considered part of the MDP; however, as deep r…