1 paper
Jingyi Zhang, Cheng Mao, Debankur Mukherjee
For stochastic gradient descent (SGD) with a constant stepsize α, the invariant law of the iterates, centered at a minimizer, describes the behavior of the algorithm over long ti…