19 citations · 36 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 3 cited
What Happens During Finetuning of Vision Transformers: An Invariance Based Investigation
Gabriele Merlin, Vedant Nanda, Ruchit Rawal +1
The pretrain-finetune paradigm usually improves downstream performance over training a model from scratch on the same task, becoming commonplace across many areas of machine learni…
cs.CL2022★ 14 cited
Language models and brains align due to more than next-word prediction and word-level information
Gabriele Merlin, Mariya Toneva
Pretrained language models have been shown to significantly predict brain recordings of people comprehending language. Recent work suggests that the prediction of the next word is…
cs.LG2022★ 19 cited
Practical Recommendations for Replay-based Continual Learning Methods
Gabriele Merlin, Vincenzo Lomonaco, Andrea Cossu +2
Continual Learning requires the model to learn from a stream of dynamic, non-stationary data without forgetting previous knowledge. Several approaches have been developed in the li…