4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2021★ 4 cited
Searching for Fast Model Families on Datacenter Accelerators
Sheng Li, Mingxing Tan, Ruoming Pang +4
Neural Architecture Search (NAS), together with model scaling, has shown remarkable progress in designing high accuracy and fast convolutional architecture families. However, as ne…
cs.LG2020
Highly Available Data Parallel ML training on Mesh Networks
Sameer Kumar, Norm Jouppi
Data parallel ML models can take several days or weeks to train on several accelerators. The long duration of training relies on the cluster of resources to be available for the jo…