Gaussian Universality of Perceptrons with Random Labels
arXiv:2205.13303 · doi:10.1103/PhysRevE.109.034305
Abstract
While classical in many theoretical settings - and in particular in statistical physics-inspired works - the assumption of Gaussian i.i.d. input data is often perceived as a strong limitation in the context of statistics and machine learning. In this study, we redeem this line of work in the case of generalized linear classification, a.k.a. the perceptron model, with random labels. We argue that there is a large universality class of high-dimensional input data for which we obtain the same minimum training loss as for Gaussian data with corresponding data covariance. In the limit of vanishing regularization, we further demonstrate that the training loss is independent of the data covariance. On the theoretical side, we prove this universality for an arbitrary mixture of homogeneous Gaussian clouds. Empirically, we show that the universality holds also for a broad range of real datasets.
References in corpus (11)
- Fashion-MNIST: a Novel Image Dataset for Benchmarking Machine Learning Algorithms
- Prevalence of Neural Collapse during the terminal phase of deep learning training
- Exploring Generalization in Deep Learning
- Data-driven emergence of convolutional structure in neural networks
- Universality of empirical risk minimization
- Kernel Alignment Risk Estimator: Risk Prediction from Training Data
- Phase Transitions in Transfer Learning for High-Dimensional Perceptrons
- High-dimensional Asymptotics of Feature Learning: How One Gradient Step Improves the Representation
- Universality in Learning from Linear Measurements
- Algorithmic pure states for the negative spherical perceptron
- Failure and success of the spectral bias prediction for Kernel Ridge Regression: the case of low-dimensional data