2 papers
cs.LG2026
Minor First, Major Last: A Depth-Induced Implicit Bias of Sharpness-Aware Minimization
Chaewon Moon, Dongkuk Si, Chulhee Yun
We study the implicit bias of Sharpness-Aware Minimization (SAM) when training -layer linear diagonal networks on linearly separable binary classification. For linear models ($L…
cs.LG2025
The Cost of Robustness: Tighter Bounds on Parameter Complexity for Robust Memorization in ReLU Nets
Yujun Kim, Chaewon Moon, Chulhee Yun
We study the parameter complexity of robust memorization for networks: the number of parameters required to interpolate any given dataset with -separation betwe…