2 papers
cs.LG2026
Knowledge Distillation from Large Reasoning Models to Compact Student Models: A Case Study on the John O Bryan Mathematics Competition
Gaurab Baral, Aaditya Khanal, Yangyang Tao +1
This paper investigates knowledge distillation from a large reasoning model (DeepSeek-R1) to a compact student model (Qwen2.5-7B). Using historical problems from the John O'Bryan M…
cs.AI2026
Beyond pass@1: A Reliability Science Framework for Long-Horizon LLM Agents
Aaditya Khanal, Yangyang Tao, Junxiu Zhou
Existing benchmarks measure capability -- whether a model succeeds on a single attempt -- but production deployments require reliability -- consistent success across repeated attem…