Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Stratos: An End-to-End Distillation Pipeline for Customized LLMs under Distributed Cloud Environments
Ziming Dai, Tuo Zhang, Fei Gao +5
The growing industrial demand for customized and cost-efficient large language models (LLMs) is fueled by the rise of vertical, domain-specific tasks and the need to optimize perfo…
cs.LG2025
HRS: Hybrid Representation Framework with Scheduling Awareness for Time Series Forecasting in Crowdsourced Cloud-Edge Platforms
Tiancheng Zhang, Cheng Zhang, Shuren Liu +3
With the rapid proliferation of streaming services, network load exhibits highly time-varying and bursty behavior, posing serious challenges for maintaining Quality of Service (QoS…