1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Xin Guo, Anran Hu, Junzi Zhang
When designing algorithms for finite-time-horizon episodic reinforcement learning problems, a common approach is to introduce a fictitious discount factor and use stationary polici…