2 papers
cs.NI2026
ImpactHO: Importance-Aware KV Cache Transfer for Multi-User Edge LLM Handover
Minwoo Kim, Soochang Song, Namyoon Lee +2
Edge LLMs must preserve inference continuity when a user hands over between edge nodes, requiring key-value (KV) cache transfer to the target node. However, simultaneous handovers…
cs.CV2026
Collaborative Edge-to-Server Inference for Vision-Language Models
Soochang Song, Yongjune Kim
We propose a collaborative edge-to-server inference framework for vision-language models (VLMs) that reduces communication cost while maintaining inference accuracy. In typical dep…