1 paper · 1 filter
Aditya Devarakonda, Ramakrishnan Kannan
Distributed-memory implementations of numerical optimization algorithm, such as stochastic gradient descent (SGD), require interprocessor communication at every iteration of the al…