2 papers
cs.LG2026
Training Specialist Models without Reasoning Trajectories for Domain Expert Distillation
Yilei Tu, Zihao Li, Shaoxiong Ji +2
Specialist distillation effectively transfers domain expertise to student models via teacher-generated reasoning trajectories. However, when these specialists are trained solely on…
cs.AI2026
Beyond Query Memorization: Large Language Model Routing with Query Decomposition and Historical Matching
Bo Lv, Jingbo Sun
Optimizing the trade-off among predictive performance and computational cost is a central focus in the deployment of Large Language Models (LLMs). Current routing methods primarily…