1 citations · 1 across the 4 of their papers we have counts for
4 papers · 1 filter
Effective Quantization of Muon Optimizer States
Aman Gupta, Rafael Celente, Abhishek Shivanna +7
The Muon optimizer, based on matrix orthogonalization, has recently shown faster convergence and better computational efficiency over AdamW in LLM pre-training. However, the memory…
Logit Attenuating Weight Normalization
Aman Gupta, Rohan Ramanath, Jun Shi +4
Over-parameterized deep networks trained using gradient-based optimizers are a popular choice for solving classification and ranking problems. Without appropriately tuned …
Lambda Learner: Fast Incremental Learning on Data Streams
Rohan Ramanath, Konstantin Salomatin, Jeffrey D. Gee +5
One of the most well-established applications of machine learning is in deciding what content to show website visitors. When observation data comes from high-velocity, user-generat…
Towards Deep and Representation Learning for Talent Search at LinkedIn
Rohan Ramanath, Hakan Inan, Gungor Polatkan +6
Talent search and recommendation systems at LinkedIn strive to match the potential candidates to the hiring needs of a recruiter or a hiring manager expressed in terms of a search…