1 paper · 1 filter
Mengtian Yang, Zhekun Zhang, Mingheng Wu +3
Deploying large-scale LLM training and inference with optimal performance is exceptionally challenging due to a complex design space of parallelism strategies, system optimizations…