21 citations · 68 across the 27 of their papers we have counts for
1 paper · 1 filter
Seanie Lee, Minki Kang, Juho Lee +2
Pre-training a large transformer model on a massive amount of unlabeled data and fine-tuning it on labeled datasets for diverse downstream tasks has proven to be a successful strat…