538 citations · 538 across the 2 of their papers we have counts for
2 papers
cs.AI2016
Efficient iterative policy optimization
Nicolas Le Roux
We tackle the issue of finding a good policy when the number of policy updates is limited. This is done by approximating the expected policy reward as a sequence of concave lower b…
math.OC2012★ 538 cited
A Stochastic Gradient Method with an Exponential Convergence Rate for Finite Training Sets
Nicolas Le Roux, Mark Schmidt, Francis Bach
We propose a new stochastic gradient method for optimizing the sum of a finite set of smooth functions, where the sum is strongly convex. While standard stochastic gradient methods…