collaborators

5 papers

cs.LG2026

Cost-Optimal LLM Routing with Limited User Feedback under User Satisfaction Guarantees

Herbert Woisetschläger, Arastun Mammadli, Ryan Zhang +1

Inference costs for large language model (LLM) applications are rapidly growing, driven by surging demand and rising infrastructure cost. Users expect high-quality responses, and i…

cs.AI2026

ActionNex: A Virtual Outage Manager for Cloud Computing

Zhenfeng Lin, Haoji Hu, Ming Hao +11

Outage management in large-scale cloud operations remains heavily manual, requiring rapid triage, cross-team coordination, and experience-driven decisions under partial observabili…

cs.LG2025

MESS+: Dynamically Learned Inference-Time LLM Routing in Model Zoos with Service Level Guarantees

Herbert Woisetschläger, Ryan Zhang, Shiqiang Wang +1

Open-weight large language model (LLM) zoos provide access to numerous high-quality models, but selecting the appropriate model for specific tasks remains challenging and requires…

cs.AI2025

SIGN: Schema-Induced Games for Naming

Ryan Zhang, Herbert Woisetschläger

Real-world AI systems are tackling increasingly complex problems, often through interactions among large language model (LLM) agents. When these agents develop inconsistent convent…

cs.AI2025

Do Students Rely on AI? Analysis of Student-ChatGPT Conversations from a Field Study

Jiayu Zheng, Lingxin Hao, Kelun Lu +13

This study explores how college students interact with generative AI (ChatGPT-4) during educational quizzes, focusing on reliance and predictors of AI adoption. Conducted at the ea…