3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.DS2026
Competitive Non-Clairvoyant KV-Cache Scheduling for LLM Inference
Yiding Feng, Zonghan Yang, Yuhao Zhang
Large Language Model (LLM) inference presents a unique scheduling challenge due to the Key-Value (KV) cache, where a job's memory footprint grows linearly with the number of decode…
cs.CL2026
Robust Uncertainty Quantification for Factual Generation of Large Language Models
Yuhao Zhang, Zhongliang Yang, Linna Zhou
The rapid advancement of large language model(LLM) technology has facilitated its integration into various domains of professional and daily life. However, the persistent challenge…
cs.AI2023★ 3 cited
TrainerAgent: Customizable and Efficient Model Training through LLM-Powered Multi-Agent System
Haoyuan Li, Hao Jiang, Tianke Zhang +6
Training AI models has always been challenging, especially when there is a need for custom models to provide personalized services. Algorithm engineers often face a lengthy process…