5 papers
Cost-Optimal LLM Routing with Limited User Feedback under User Satisfaction Guarantees
Herbert Woisetschläger, Arastun Mammadli, Ryan Zhang +1
Inference costs for large language model (LLM) applications are rapidly growing, driven by surging demand and rising infrastructure cost. Users expect high-quality responses, and i…
ActionNex: A Virtual Outage Manager for Cloud Computing
Zhenfeng Lin, Haoji Hu, Ming Hao +11
Outage management in large-scale cloud operations remains heavily manual, requiring rapid triage, cross-team coordination, and experience-driven decisions under partial observabili…
MESS+: Dynamically Learned Inference-Time LLM Routing in Model Zoos with Service Level Guarantees
Herbert Woisetschläger, Ryan Zhang, Shiqiang Wang +1
Open-weight large language model (LLM) zoos provide access to numerous high-quality models, but selecting the appropriate model for specific tasks remains challenging and requires…
SIGN: Schema-Induced Games for Naming
Ryan Zhang, Herbert Woisetschläger
Real-world AI systems are tackling increasingly complex problems, often through interactions among large language model (LLM) agents. When these agents develop inconsistent convent…
Do Students Rely on AI? Analysis of Student-ChatGPT Conversations from a Field Study
Jiayu Zheng, Lingxin Hao, Kelun Lu +13
This study explores how college students interact with generative AI (ChatGPT-4) during educational quizzes, focusing on reliance and predictors of AI adoption. Conducted at the ea…