3 papers
cs.DC2026
CALVO: Improve Serving Efficiency for LLM Inferences with Intense Network Demands
Weiye Wang, Chen Chen, Junxue Zhang +7
Distributed prefix caching has become a core technique for efficient LLM serving. However, for long-context requests with high cache hit ratios, retrieving reusable KVCache blocks…
cs.CV2024
Perceptual Quality Assessment of Trisoup-Lifting Encoded 3D Point Clouds
Juncheng Long, Honglei Su, Qi Liu +4
No-reference bitstream-layer point cloud quality assessment (PCQA) can be deployed without full decoding at any network node to achieve real-time quality monitoring. In this work,…
cs.MM2024
Perceptual Quality Assessment of Octree-RAHT Encoded 3D Point Clouds
Dongshuai Duan, Honglei Su, Qi Liu +4
No-reference bitstream-layer point cloud quality assessment (PCQA) can be deployed without full decoding at any network node to achieve real-time quality monitoring. In this work,…