5 citations · 6 across the 2 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2021
Benchmarking down-scaled (not so large) pre-trained language models
M. Aßenmacher, P. Schulze, C. Heumann
Large Transformer-based language models are pre-trained on corpora of varying sizes, for a different number of steps and with different batch sizes. At the same time, more fundamen…
cs.CL2021
Re-Evaluating GermEval17 Using German Pre-Trained Language Models
M. Aßenmacher, A. Corvonato, C. Heumann
The lack of a commonly used benchmark data set (collection) such as (Super-)GLUE (Wang et al., 2018, 2019) for the evaluation of non-English pre-trained language models is a severe…