2 papers
cs.NI2025
Quality-of-Service Aware LLM Routing for Edge Computing with Multiple Experts
Jin Yang, Qiong Wu, Zhiying Feng +3
Large Language Models (LLMs) have demonstrated remarkable capabilities, leading to a significant increase in user demand for LLM services. However, cloud-based LLM services often s…
cs.LG2023
FedDD: Toward Communication-efficient Federated Learning with Differential Parameter Dropout
Zhiying Feng, Xu Chen, Qiong Wu +3
Federated Learning (FL) requires frequent exchange of model parameters, which leads to long communication delay, especially when the network environments of clients vary greatly. M…