571 citations · 573 across the 2 of their papers we have counts for
2 papers
cs.LG2016★ 571 cited
On Large-Batch Training for Deep Learning: Generalization Gap and Sharp Minima
Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal +2
The stochastic gradient descent (SGD) method and its variants are algorithms of choice for many Deep Learning tasks. These methods operate in a small-batch regime wherein a fractio…
cs.CE2014★ 2 cited
Identification of Helicopter Dynamics based on Flight Data using Nature Inspired Techniques
S. N. Omkar, Dheevatsa Mudigere, J Senthilnath +1
The complexity of helicopter flight dynamics makes modeling and helicopter system identification a very difficult task. Most of the traditional techniques require a model structure…