2 papers
cs.DC2025
FREESH: Fair, Resource- and Energy-Efficient Scheduling for LLM Serving on Heterogeneous GPUs
Xuan He, Zequan Fang, Jinzhao Lian +3
The ever-increasing computation and energy demand for LLM and AI agents call for holistic and efficient optimization of LLM serving systems. In practice, heterogeneous GPU clusters…
eess.SY2025
Voltage Regulation in Distribution Systems with Data Center Loads
Yize Chen, Baosen Zhang
Recent boom in foundation models and AI computing have raised growing concerns on the power and energy trajectories of large-scale data centers. This paper focuses on the voltage i…