2 citations · 2 across the 2 of their papers we have counts for
2 papers
stat.ML2024★ 2 cited
Position: Understanding LLMs Requires More Than Statistical Generalization
Patrik Reizinger, Szilvia Ujváry, Anna Mészáros +3
The last decade has seen blossoming research in deep learning theory attempting to answer, "Why does deep learning generalize?" A powerful shift in perspective precipitated this pr…
stat.ML2022
Rethinking Sharpness-Aware Minimization as Variational Inference
Szilvia Ujváry, Zsigmond Telek, Anna Kerekes +2
Sharpness-aware minimization (SAM) aims to improve the generalisation of gradient-based learning by seeking out flat minima. In this work, we establish connections between SAM and…