7 citations · 8 across the 7 of their papers we have counts for
4 papers · 1 filter
Forget Sharpness: Perturbed Forgetting of Model Biases Within SAM Dynamics
Ankit Vani, Frederick Tung, Gabriel L. Oliveira +1
Despite attaining high empirical generalization, the sharpness of models trained with sharpness-aware minimization (SAM) do not always correlate with generalization error. Instead…
Pretext Training Algorithms for Event Sequence Data
Yimu Wang, He Zhao, Ruizhi Deng +2
Pretext training followed by task-specific fine-tuning has been a successful approach in vision and language domains. This paper proposes a self-supervised pretext training framewo…
AdaFlood: Adaptive Flood Regularization
Wonho Bae, Yi Ren, Mohamad Osama Ahmed +3
Although neural networks are conventionally optimized towards zero training loss, it has been recently learned that targeting a non-zero training loss threshold, referred to as a f…
Meta Temporal Point Processes
Wonho Bae, Mohamed Osama Ahmed, Frederick Tung +1
A temporal point process (TPP) is a stochastic process where its realization is a sequence of discrete events in time. Recent work in TPPs model the process using a neural network…