3 papers
cs.DC2026
Coral: Cost-Efficient Multi-LLM Serving over Heterogeneous Cloud GPUs
Yixuan Mei, Zikun Li, Zixuan Chen +5
The usage of large language models (LLMs) has grown increasingly fragmented, with no single model dominating. Meanwhile, cloud providers offer a wide range of mid-tier and older-ge…
cs.IT2026
Bandwidth Cost of Locally Repairable Convertible Codes in the Global Merge Regime
Saransh Chopra, Shubhransh Singhvi, K. V. Rashmi
Recent studies have shown that distributed storage systems can achieve significant space savings by adapting redundancy levels to varying disk failure rates. This adaptation is per…
cs.IT2025
Tight Lower Bounds on the Bandwidth Cost of MDS Convertible Codes in the Split Regime
Shubhransh Singhvi, Saransh Chopra, K. V. Rashmi
Recent advances in erasure coding for distributed storage systems have demonstrated that adapting redundancy to varying disk failure rates can lead to substantial storage savings.…