3 papers
cs.LG2025
Stratos: An End-to-End Distillation Pipeline for Customized LLMs under Distributed Cloud Environments
Ziming Dai, Tuo Zhang, Fei Gao +5
The growing industrial demand for customized and cost-efficient large language models (LLMs) is fueled by the rise of vertical, domain-specific tasks and the need to optimize perfo…
cs.LG2025
HRS: Hybrid Representation Framework with Scheduling Awareness for Time Series Forecasting in Crowdsourced Cloud-Edge Platforms
Tiancheng Zhang, Cheng Zhang, Shuren Liu +3
With the rapid proliferation of streaming services, network load exhibits highly time-varying and bursty behavior, posing serious challenges for maintaining Quality of Service (QoS…
cs.NI2022
EdgeMatrix: A Resources Redefined Edge-Cloud System for Prioritized Services
Yuanming Ren, Shihao Shen, Yanli Ju +3
The edge-cloud system has the potential to combine the advantages of heterogeneous devices and truly realize ubiquitous computing. However, for service providers to guarantee the S…