Showing cs.DCShow all
3 papers · 1 filter
cs.DC2026
Beyond Microservices: Testing Web-Scale RCA Methods on GPU-Driven LLM Workloads
Dominik Scheinert, Alexander Acker, Thorsten Wittkopp +6
Large language model (LLM) services have become an integral part of search, assistance, and decision-making applications. However, unlike traditional web or microservices, the hard…
cs.DC2026
Distributed LLM Pretraining During Renewable Curtailment Windows: A Feasibility Study
Philipp Wiesner, Soeren Becker, Brett Cornick +3
Training large language models (LLMs) requires substantial compute and energy. At the same time, renewable energy sources regularly produce more electricity than the grid can absor…
cs.DC2025
What happens when nanochat meets DiLoCo?
Alexander Acker, Soeren Becker, Sasho Nedelkoski +3
Although LLM training is typically centralized with high-bandwidth interconnects and large compute budgets, emerging methods target communication-constrained training in distribute…