2 papers
cs.AR2025
CLONE: Customizing LLMs for Efficient Latency-Aware Inference at the Edge
Chunlin Tian, Xinpeng Qin, Kahou Tam +5
Deploying large language models (LLMs) on edge devices is crucial for delivering fast responses and ensuring data privacy. However, the limited storage, weight, and power of edge d…
cs.LG2024
Ranking-based Client Selection with Imitation Learning for Efficient Federated Learning
Chunlin Tian, Zhan Shi, Xinpeng Qin +2
Federated Learning (FL) enables multiple devices to collaboratively train a shared model while ensuring data privacy. The selection of participating devices in each training round…