knowledge distillation 1model calibration 1multi-teacher training 1on-policy distillation 1over-calling 1tool use 1
From the 1 of 5 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
When Top-K Misses the Decision: Tool-Call Drift in Multi-Teacher On-Policy Distillation
Jiabin Shen, Guang Chen, Chengjun Mao
The paper studies how multi-teacher on-policy distillation can cause language models to over-call tools, and introduces Soft Clamp, a token-level divergence calibration method that…
cs.CL2026
Attention-guided Evidence Grounding for Spoken Question Answering
Ke Yang, Bolin Chen, Yuejie Li +5
Spoken Question Answering (Spoken QA) presents a challenging cross-modal problem: effectively aligning acoustic queries with textual knowledge while avoiding the latency and error…