1 paper
Mingyuan Wang, Yangzi Guo, Sida Liu +1
The remarkable performance of modern deep neural networks (DNNs) is largely driven by their massive scale, often comprising tens to hundreds of millions-or even billions-of paramet…