1 paper
Viktoriia Chekalina, Anna Rudenko, Gleb Mezentsev +3
The performance of Transformer models has been enhanced by increasing the number of parameters and the length of the processed text. Consequently, fine-tuning the entire model beco…