Showing cs.DCShow all
2 papers · 1 filter
cs.DC2026
DUAL-BLADE: Dual-Path NVMe-Direct KV-Cache Offloading for Edge LLM Inference
Bodon Jeong, Hongsu Byun, Youngjae Kim +4
The increasing deployment of Large Language Model (LLM) inference on edge AI systems demands efficient execution under tight memory budgets. A key challenge arises from Key-Value (…
cs.DC2026
AFLL: Real-time Load Stabilization for MMO Game Servers Based on Circular Causality Learning
Shinsuk Kang, Youngjae Kim
Massively Multiplayer Online (MMO) game servers must handle thousands of simultaneous players while maintaining sub-100ms response times. When server load exceeds capacity, traditi…