4 papers
Low-Latency Edge LLM Handover via Joint KV Cache Transfer and Token Prefill
Seunghun Lee, Jihong Park, Ce Zheng +1
Edge deployment of large language models (LLMs) can reduce latency for interactive services, but mobility introduces service interruptions when an user equipment (UE) hands over be…
Low-Complexity Semantic Packet Aggregation for Token Communication via Lookahead Search
Seunghun Lee, Jihong Park, Jinho Choi +1
Tokens are fundamental processing units of generative AI (GenAI) and large language models (LLMs), and token communication (TC) is essential for enabling remote AI-generate content…
Semantic Packet Aggregation for Token Communication via Genetic Beam Search
Seunghun Lee, Jihong Park, Jinho Choi +1
Token communication (TC) is poised to play a pivotal role in emerging language-driven applications such as AI-generated content (AIGC) and wireless language models (LLMs). However,…
Semantic Packet Aggregation and Repeated Transmission for Text-to-Image Generation
Seunghun Lee, Jihong Park, Jinho Choi +1
Text-based communication is expected to be prevalent in 6G applications such as wireless AI-generated content (AIGC). Motivated by this, this paper addresses the challenges of tran…