1 paper
Mengtian Yang, Zhekun Zhang, Mingheng Wu +3
Deploying large-scale LLM training and inference with optimal performance is exceptionally challenging due to a complex design space of parallelism strategies, system optimizations…