5 citations · 5 across the 2 of their papers we have counts for
1 paper · 1 filter
Lukas Gienapp, Christopher Schröder, Stefan Schweter +5
Large language model development relies on large-scale training corpora, yet most contain data of unclear licensing status, limiting the development of truly open models. This prob…