1 paper · 1 filter
Hadi Pouransari, Chun-Liang Li, Jen-Hao Rick Chang +4
Large language models (LLMs) are commonly trained on datasets consisting of fixed-length token sequences. These datasets are created by randomly concatenating documents of various…