Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Sparse Linear Regression and Lattice Problems
Aparna Gupte, Neekon Vafa, Vinod Vaikuntanathan
Sparse linear regression (SLR) is a well-studied problem in statistics where one is given a design matrix and a response vector for a -s…
cs.LG2024
SGD and Weight Decay Secretly Minimize the Rank of Your Neural Network
Tomer Galanti, Zachary S. Siegel, Aparna Gupte +1
We investigate the inherent bias of Stochastic Gradient Descent (SGD) toward learning low-rank weight matrices during the training of deep neural networks. Our results demonstrate…