1 paper · 1 filter
Shuyao Xu, Cheng Peng, Jiangxuan Long +3
Recent advances in model distillation show that data from advanced reasoning models can effectively train smaller student models. However, standard practices discard incorrect reas…