48 citations · 48 across the 2 of their papers we have counts for
1 paper · 1 filter
Xia Han, Ruodu Wang, Xun Yu Zhou
We propose \emph{Choquet regularizers} to measure and manage the level of exploration for reinforcement learning (RL), and reformulate the continuous-time entropy-regularized RL pr…