3 citations · 14 across the 12 of their papers we have counts for
1 paper · 1 filter
Noufel Frikha, Maximilien Germain, Mathieu Laurière +2
We study policy gradient for mean-field control in continuous time in a reinforcement learning setting. By considering randomised policies with entropy regularisation, we derive a…