25 citations · 32 across the 11 of their papers we have counts for
1 paper · 1 filter
Gabriele Merlin, Vedant Nanda, Ruchit Rawal +1
The pretrain-finetune paradigm usually improves downstream performance over training a model from scratch on the same task, becoming commonplace across many areas of machine learni…