113 citations · 113 across the 1 of their papers we have counts for
1 paper
Tanja Bunk, Daksh Varshneya, Vladimir Vlasov +1
Large-scale pre-trained language models have shown impressive results on language understanding benchmarks like GLUE and SuperGLUE, improving considerably over other pre-training m…