2 papers
cs.LG2026
Smoothing the Score Function for Generalization in Diffusion Models: An Optimization-based Explanation Framework
Xinyu Zhou, Jiawei Zhang, Stephen J. Wright
Diffusion models achieve remarkable generation quality, yet face a fundamental challenge known as memorization, where generated samples can replicate training samples exactly. We d…
math.OC2025
First-ish Order Methods: Hessian-aware Scalings of Gradient Descent
Oscar Smee, Fred Roosta, Stephen J. Wright
Gradient descent is the primary workhorse for optimizing large-scale problems in machine learning. However, its performance is highly sensitive to the choice of the learning rate.…