1 paper
Zhuo Chen, Yuxuan Miao, Supryadi +1
Large language models (LLMs) rely on pretraining on massive and heterogeneous corpora, where training data composition has a decisive impact on training efficiency and downstream g…