3 papers
cs.IR2025
Time Matters: Enhancing Sequential Recommendations with Time-Guided Graph Neural ODEs
Haoyan Fu, Zhida Qin, Shixiao Yang +5
Sequential recommendation (SR) is widely deployed in e-commerce platforms, streaming services, etc., revealing significant potential to enhance user experience. However, existing m…
cs.LG2025
DiffKV: Differentiated Memory Management for Large Language Models with Parallel KV Compaction
Yanqi Zhang, Yuwei Hu, Runyuan Zhao +2
Large language models (LLMs) demonstrate remarkable capabilities but face substantial serving costs due to their high memory demands, with the key-value (KV) cache being a primary…
cs.NI2024
cRVR: A Stackelberg Game Approach for Joint Privacy-Aware Video Requesting and Edge Caching
Xianzhi Zhang, Linchang Xiao, Yipeng Zhou +4
As users conveniently stream their favorite online videos, video request records are automatically stored by video content providers, which have a high chance of privacy leakage. U…