Showing cs.DCShow all
3 papers · 1 filter
cs.DC2026
Concordia: JIT-Compiled Persistent-Kernel Checkpointing for Fault-Tolerant LLM Inference
Yuhang Gan, Yiwei Yang, Yuyi Li +6
Long-running LLM agents keep valuable state resident on GPUs: KV caches, request schedulers, communication state, and sometimes online adapters. Losing this state after a GPU or co…
cs.DC2026
PlanetServe: A Decentralized, Scalable, and Privacy-Preserving Overlay for Democratizing Large Language Model Serving
Fei Fang, Yifan Hua, Shengze Wang +4
While significant progress has been made in research and development on open-source and cost-efficient large-language models (LLMs), serving scalability remains a critical challeng…
cs.DC2025
CloudQC: A Network-aware Framework for Multi-tenant Distributed Quantum Computing
Ruilin Zhou, Yuhang Gan, Yi Liu +1
Distributed quantum computing (DQC) that allows a large quantum circuit to be executed simultaneously on multiple quantum processing units (QPUs) becomes a promising approach to in…