2 papers
cs.DC2025
Optimizing Distributed Deployment of Mixture-of-Experts Model Inference in Serverless Computing
Mengfan Liu, Wei Wang, Chuan Wu
With the advancement of serverless computing, running machine learning (ML) inference services over a serverless platform has been advocated, given its labor-free scalability and c…
cs.LG2024
Echo: Simulating Distributed Training At Scale
Yicheng Feng, Yuetao Chen, Kaiwen Chen +7
Simulation offers unique values for both enumeration and extrapolation purposes, and is becoming increasingly important for managing the massive machine learning (ML) clusters and…