3 papers
cs.LG2026
DP-Muon: Differentially Private Optimization via Matrix-Orthogonalized Momentum
Jihwan Kim, Chenglin Fan
We study differentially private (DP) training with Muon, a matrix-valued optimizer that updates hidden-layer weights using momentum followed by Newton--Schulz orthogonalization. Wh…
cs.LG2026
Robust and Consistent Ski Rental with Distributional Advice
Jihwan Kim, Chenglin Fan
The ski rental problem is a canonical model for online decision-making under uncertainty, capturing the fundamental trade-off between repeated rental costs and a one-time purchase.…
cs.LG2026
Scaling Laws of SignSGD in Linear Regression: When Does It Outperform SGD?
Jihwan Kim, Dogyoon Song, Chulhee Yun
We study scaling laws of signSGD under a power-law random features (PLRF) model that accounts for both feature and target decay. We analyze the population risk of a linear model tr…