1 paper · 1 filter
Muyu He, Muhammad Ali Shafique, Anand Kumar +2
Distilling the thinking traces of a Large Language Model (LLM) with reasoning capabilities into a smaller model has been proven effective. Yet, there is a scarcity of work done on…