1 paper
Masato Mita, Taiga Someya, Ryo Yoshida +1
Large Language Models (LLMs) remain substantially less data-efficient than humans. Pre-pretraining (PPT) on synthetic languages has been proposed to close this gap, with prior work…