1 paper · 1 filter
Srikrishna Iyer
We present our submission to the BabyLM challenge, aiming to push the boundaries of data-efficient language model pretraining. Our method builds upon deep mutual learning, introduc…