49 citations · 50 across the 2 of their papers we have counts for
4 papers
Nonlinear Initialization Methods for Low-Rank Neural Networks
Kiran Vodrahalli, Rakesh Shivanna, Maheswaran Sathiamoorthy +2
We propose a novel low-rank initialization framework for training low-rank deep neural networks -- networks where the weight parameters are re-parameterized by products of two low-…
Understanding and Improving Knowledge Distillation
Jiaxi Tang, Rakesh Shivanna, Zhe Zhao +4
Knowledge Distillation (KD) is a model-agnostic technique to improve model quality while having a fixed capacity budget. It is a commonly used technique for model compression, wher…
Towards Neural Mixture Recommender for Long Range Dependent User Sequences
Jiaxi Tang, Francois Belletti, Sagar Jain +4
Understanding temporal dynamics has proved to be highly valuable for accurate recommendation. Sequential recommenders have been successful in modeling the dynamics of users and ite…
Seq2Slate: Re-ranking and Slate Optimization with RNNs
Irwan Bello, Sayali Kulkarni, Sagar Jain +6
Ranking is a central task in machine learning and information retrieval. In this task, it is especially important to present the user with a slate of items that is appealing as a w…