2 papers
cs.LG2026
Scaling Federated Linear Contextual Bandits via Sketching
Hantao Yang, Hong Xie, Xutong Liu +1
In federated contextual linear bandits, high data dimensionality incurs prohibitive computation and communication costs: local agents perform -time determinant computation…
cs.CL2025
LLM Cache Bandit Revisited: Addressing Query Heterogeneity for Cost-Effective LLM Inference
Hantao Yang, Hong Xie, Defu Lian +1
This paper revisits the LLM cache bandit problem, with a special focus on addressing the query heterogeneity for cost-effective LLM inference. Previous works often assume uniform q…