1 paper
Chandrashekar Lakshminarayanan, Amit Vikram Singh
Understanding the role of (stochastic) gradient descent (SGD) in the training and generalisation of deep neural networks (DNNs) with ReLU activation has been the object study in th…