6 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.PF2026
Measured Joules, Learned Routes: Learning to Route for Energy-Efficient LLM Serving
Muhammad Abdur Rab Siddiqui, Daniela Rojas, Chen Yang +3
Large language models (LLMs) and agentic AI systems are creating rapidly growing inference energy demands as model sizes grow and reasoning trajectories extend. While in practice,…
cs.AI2026
Beyond Imitation: Filtering On-Policy Distillation by Reasoning Progress
Chen Yang, Haiyuan Wan, Rengrong Xiong +2
On-policy distillation (OPD) has emerged as an effective framework for post-training language models by pairing student-generated trajectories with dense token-level supervision fr…
cs.AI2024★ 6 cited
Long Term Memory: The Foundation of AI Self-Evolution
Xun Jiang, Feng Li, Han Zhao +12
Large language models (LLMs) like GPTs, trained on vast datasets, have demonstrated impressive capabilities in language understanding, reasoning, and planning, achieving human-leve…