22 citations · 65 across the 9 of their papers we have counts for
1 paper · 1 filter
Zhimin Hou, Kuangen Zhang, Yi Wan +3
The optimal policy of a reinforcement learning problem is often discontinuous and non-smooth. I.e., for two states with similar representations, their optimal policies can be signi…