26 citations · 26 across the 2 of their papers we have counts for
1 paper · 1 filter
Maxim Kaledin, Eric Moulines, Alexey Naumov +2
Linear two-timescale stochastic approximation (SA) scheme is an important class of algorithms which has become popular in reinforcement learning (RL), particularly for the policy e…