3 papers
cs.CL2026
AdaSpec: Adaptive Speculative Decoding for Fast, SLO-Aware Large Language Model Serving
Kaiyu Huang, Hao Wu, Zhubo Shi +3
Cloud-based Large Language Model (LLM) services often face challenges in achieving low inference latency and meeting Service Level Objectives (SLOs) under dynamic request patterns.…
eess.SP2025
Exactly or Approximately Wasserstein Distributionally Robust Estimation According to Wasserstein Radii Being Small or Large
Xiao Ding, Enbin Song, Dunbiao Niu +2
This paper primarily considers the robust estimation problem under Wasserstein distance constraints on the parameter and noise distributions in the linear measurement model with ad…
eess.SP2024
Coordinated Spectral Efficiency Prediction for Real-World 5G CoMP Systems
Zhixing Chen, Zhaoyu Fan, Yang Li +3
Coordinated multipoint (CoMP) systems incur substantial resource consumption due to the management of backhaul links and the coordination among various base stations (BSs). Accurat…