1 paper · 1 filter
Hritik Bansal, Arian Hosseini, Rishabh Agarwal +2
Training on high-quality synthetic data from strong language models (LMs) is a common strategy to improve the reasoning performance of LMs. In this work, we revisit whether this st…