10 citations · 10 across the 3 of their papers we have counts for
1 paper · 1 filter
Zhimin Hou, Kuangen Zhang, Yi Wan +3
The optimal policy of a reinforcement learning problem is often discontinuous and non-smooth. I.e., for two states with similar representations, their optimal policies can be signi…