Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Generalization, memorization, and overfitting for diffusion models trained in the lazy high-dimensional regime
Hugo Latourelle-Vigeant, Sinho Chewi, Aram-Alexandre Pooladian +2
Modern score-based generative models have achieved remarkable empirical success in high-dimensional tasks such as image, audio, and video synthesis. These models reduce distributio…
stat.ML2026
Spectral Lens: Activation and Gradient Spectra as Diagnostics of LLM Optimization
Andy Zeyi Liu, Elliot Paquette, John Sous
Training loss and throughput can hide distinct internal representation in language-model training. To examine these hidden mechanics, we use spectral measurements as practical and…