2 citations · 2 across the 1 of their papers we have counts for
1 paper
Noufel Frikha, Maximilien Germain, Mathieu Laurière +2
We study policy gradient for mean-field control in continuous time in a reinforcement learning setting. By considering randomised policies with entropy regularisation, we derive a…