2 papers
cs.LG2026
Cost-Optimal LLM Routing with Limited User Feedback under User Satisfaction Guarantees
Herbert Woisetschläger, Arastun Mammadli, Ryan Zhang +1
Inference costs for large language model (LLM) applications are rapidly growing, driven by surging demand and rising infrastructure cost. Users expect high-quality responses, and i…
cs.AI2026
AgentSearchBench: A Benchmark for AI Agent Search in the Wild
Bin Wu, Arastun Mammadli, Xiaoyu Zhang +1
The rapid growth of AI agent ecosystems is transforming how complex tasks are delegated and executed, creating a new challenge of identifying suitable agents for a given task. Unli…