1 paper
M. Aßenmacher, A. Corvonato, C. Heumann
The lack of a commonly used benchmark data set (collection) such as (Super-)GLUE (Wang et al., 2018, 2019) for the evaluation of non-English pre-trained language models is a severe…