9 citations · 16 across the 2 of their papers we have counts for
3 papers
cs.LG2019
Memory-Efficient Adaptive Optimization
Rohan Anil, Vineet Gupta, Tomer Koren +1
Adaptive gradient-based optimizers such as Adagrad and Adam are crucial for achieving state-of-the-art performance in machine translation and language modeling. However, these meth…
cs.LG2017★ 9 cited
A Unified Approach to Adaptive Regularization in Online and Stochastic Optimization
Vineet Gupta, Tomer Koren, Yoram Singer
We describe a framework for deriving and analyzing online optimization algorithms that incorporate adaptive, data-dependent regularization, also termed preconditioning. Such algori…
cs.LG2017★ 7 cited
Random Features for Compositional Kernels
Amit Daniely, Roy Frostig, Vineet Gupta +1
We describe and analyze a simple random feature scheme (RFS) from prescribed compositional kernels. The compositional kernels we use are inspired by the structure of convolutional…