1 paper
Cole Hawkins, Xing Liu, Zheng Zhang
While post-training model compression can greatly reduce the inference cost of a deep neural network, uncompressed training still consumes a huge amount of hardware resources, run-…