2 citations · 2 across the 1 of their papers we have counts for
1 paper
Safa Messaoud, Billel Mokeddem, Zhenghai Xue +4
Learning expressive stochastic policies instead of deterministic ones has been proposed to achieve better stability, sample complexity, and robustness. Notably, in Maximum Entropy…