53 citations · 54 across the 2 of their papers we have counts for
2 papers
cs.DC2020★ 1 cited
Scaling Distributed Training with Adaptive Summation
Saeed Maleki, Madan Musuvathi, Todd Mytkowicz +5
Stochastic gradient descent (SGD) is an inherently sequential training algorithm--computing the gradient at batch depends on the model parameters learned from batch . Prio…
cs.LG2019★ 53 cited
Machine Learning at Microsoft with ML .NET
Zeeshan Ahmed, Saeed Amizadeh, Mikhail Bilenko +31
Machine Learning is transitioning from an art and science into a technology available to every developer. In the near future, every application on every platform will incorporate t…