13 citations · 33 across the 24 of their papers we have counts for
Showing 2026 · stat.MLShow all
3 papers · 2 filters
stat.ML2026
Next-token functional estimation
Milind Nakul, Vidya Muthukumar, Ashwin Pananjady
Suppose we observe the first points of a sequence of random variables having length , and wish to estimate a functional of the unobserved final point and the empirical mea…
stat.ML2026
SGD Provably Prioritizes a Shortcut Spurious Feature in the XOR Model
Tyler LaBonte, Vidya Muthukumar
Neural networks are known to be susceptible to over-reliance on spurious correlations. However, the precise mechanism by which models exploit shortcut features is not fully underst…
stat.ML2026
How Does the ReLU Activation Affect the Implicit Bias of Gradient Descent on High-dimensional Neural Network Regression?
Kuo-Wei Lai, Guanghui Wang, Molei Tao +1
Overparameterized ML models, including neural networks, typically induce underdetermined training objectives with multiple global minima. The implicit bias refers to the limiting g…