2 citations · 2 across the 1 of their papers we have counts for
1 paper
Risto Luukkonen, Jonathan Burdge, Elaine Zosa +5
The pretraining of state-of-the-art large language models now requires trillions of words of text, which is orders of magnitude more than available for the vast majority of languag…