4 papers
Adapter-Augmented Bandits for Online Multi-Constrained Multi-Modal Inference Scheduling
Xianzhi Zhang, Yue Xu, Yinlin Zhu +4
Multi-modal large language model (MLLM) inference scheduling enables strong response quality under practical and heterogeneous budgets, beyond what a homogeneous single-backend set…
A Survey on Privacy-Preserving Caching at Network Edge: Classification, Solutions, and Challenges
Xianzhi Zhang, Yipeng Zhou, Di Wu +4
Caching content at the edge network is a popular and effective technique widely deployed to alleviate the burden of network backhaul, shorten service delay and improve service qual…
BGTplanner: Maximizing Training Accuracy for Differentially Private Federated Recommenders via Strategic Privacy Budget Allocation
Xianzhi Zhang, Yipeng Zhou, Miao Hu +4
To mitigate the rising concern about privacy leakage, the federated recommender (FR) paradigm emerges, in which decentralized clients co-train the recommendation model without expo…
cRVR: A Stackelberg Game Approach for Joint Privacy-Aware Video Requesting and Edge Caching
Xianzhi Zhang, Linchang Xiao, Yipeng Zhou +4
As users conveniently stream their favorite online videos, video request records are automatically stored by video content providers, which have a high chance of privacy leakage. U…