1 paper · 1 filter
Jingyi Zhang, Cheng Mao, Debankur Mukherjee
For stochastic gradient descent (SGD) with a constant stepsize α, the invariant law of the iterates, centered at a minimizer, describes the behavior of the algorithm over long ti…