90 citations · 165 across the 16 of their papers we have counts for
1 paper · 1 filter
Noufel Frikha, Maximilien Germain, Mathieu Laurière +2
We study policy gradient for mean-field control in continuous time in a reinforcement learning setting. By considering randomised policies with entropy regularisation, we derive a…