1 paper
Xiaoxuan Zhu, Zhouhong Gu, Baiqian Wu +5
Pre-training large language models (LLMs) necessitates enormous diverse textual corpora, making effective data selection a key challenge for balancing computational resources and m…