3 citations · 3 across the 1 of their papers we have counts for
1 paper
Bargav Jayaraman, Esha Ghosh, Melissa Chase +3
Pre-trained large language models, such as GPT\nobreakdash-2 and BERT, are often fine-tuned to achieve state-of-the-art performance on a downstream task. One natural example is the…