1 paper · 1 filter
Aditya Shah, Aditya Challa, Sravan Danda +2
Stochastic Gradient Descent (SGD) is the main approach to optimizing neural networks. Several generalization properties of deep networks, such as convergence to a flatter minima, a…