3 papers
cs.OS2026
SSV: Sparse Speculative Verification for Efficient LLM Inference
Zhibin Wang, Ziyu Zhong, Nuo Shen +3
Speculative decoding and dynamic sparse attention are two complementary approaches for accelerating long-context LLM inference: the former amortizes target-model execution across m…
cs.NI2025
Predictability-Aware Motion Prediction for Edge XR via High-Order Error-State Kalman Filtering
Ziyu Zhong, Björn Landfeldt, Günter Alce +1
As 6G networks are developed and defined, offloading of XR applications is emerging as one of the strong new use cases. The reduced 6G latency coupled with edge processing infrastr…
cs.NI2025
Video Streaming with Kairos: An MPC-Based ABR with Streaming-Aware Throughput Prediction
Ziyu Zhong, Mufan Liu, Le Yang +3
In this paper, we present Kairos, a model predictive control (MPC)-based adaptive bitrate (ABR) scheme that integrates streaming-aware throughput predictions to enhance video strea…