1 paper
Quchen Fu, Ramesh Chukka, Keith Achorn +5
GPUs have been favored for training deep learning models due to their highly parallelized architecture. As a result, most studies on training optimization focus on GPUs. There is o…