1 citations · 2 across the 8 of their papers we have counts for
1 paper · 1 filter
Xin Guo, Anran Hu, Junzi Zhang
When designing algorithms for finite-time-horizon episodic reinforcement learning problems, a common approach is to introduce a fictitious discount factor and use stationary polici…