1 paper
Neeraj Gangwar, Suma P Bhat, Nickvash Kani
While large models pre-trained on high-quality data exhibit excellent performance on mathematical reasoning (e.g., GSM8k, MultiArith), it remains challenging to specialize smaller…