collaborators

5 papers

cs.LG2026

SAGE: Training-Free Semantic Evidence Composition for Edge-Cloud Inference under Hard Uplink Budgets

Inhyeok Choi, Hyuncheol Park

Edge-cloud hybrid inference offloads difficult inputs to a powerful remote model, but the uplink channel imposes hard per-request constraints on the number of bits that can be tran…

eess.SP2026

Low-Latency Edge LLM Handover via Joint KV Cache Transfer and Token Prefill

Seunghun Lee, Jihong Park, Ce Zheng +1

Edge deployment of large language models (LLMs) can reduce latency for interactive services, but mobility introduces service interruptions when an user equipment (UE) hands over be…

eess.SP2025

Low-Complexity Semantic Packet Aggregation for Token Communication via Lookahead Search

Seunghun Lee, Jihong Park, Jinho Choi +1

Tokens are fundamental processing units of generative AI (GenAI) and large language models (LLMs), and token communication (TC) is essential for enabling remote AI-generate content…

eess.SP2025

Semantic Packet Aggregation for Token Communication via Genetic Beam Search

Seunghun Lee, Jihong Park, Jinho Choi +1

Token communication (TC) is poised to play a pivotal role in emerging language-driven applications such as AI-generated content (AIGC) and wireless language models (LLMs). However,…

eess.SP2025

Semantic Packet Aggregation and Repeated Transmission for Text-to-Image Generation

Seunghun Lee, Jihong Park, Jinho Choi +1

Text-based communication is expected to be prevalent in 6G applications such as wireless AI-generated content (AIGC). Motivated by this, this paper addresses the challenges of tran…