1 citations · 1 across the 6 of their papers we have counts for
1 paper · 2 filters
Weihao Cui, Ji Zhang, Han Zhao +5
The rapid proliferation of large language models has driven the need for efficient GPU training clusters. However, it is challenging due to the frequent occurrence of training anom…