3 citations · 3 across the 1 of their papers we have counts for
1 paper
V. D. Viellieber, M. Aßenmacher
Recently it has been shown that large pre-trained language models like BERT (Devlin et al., 2018) are able to store commonsense factual knowledge captured in its pre-training corpu…