1 paper · 1 filter
Felix Draxler, Kambis Veschgini, Manfred Salmhofer +1
Training neural networks involves finding minima of a high-dimensional non-convex loss function. Knowledge of the structure of this energy landscape is sparse. Relaxing from linear…