2 papers
cs.DC2026
OffloadFS: Leveraging Disaggregated Storage for Computation Offloading
Sungho Moon, Daegyu Han, Hera Koo +4
Disaggregated storage systems improve resource utilization and enable independent scaling of storage and compute resources by separating storage resources from computing resources…
cs.DC2025
Accelerating LLM Inference with Precomputed Query Storage
Jay H. Park, Youngju Cho, Choungsol Lee +2
Large language model (LLM) inference often suffers from high latency, particularly in resource-constrained environments such as on-device or edge deployments. To address this chall…