19 citations · 19 across the 3 of their papers we have counts for
3 papers
cs.LG2022
Nesting Forward Automatic Differentiation for Memory-Efficient Deep Neural Network Training
Cong Guo, Yuxian Qiu, Jingwen Leng +6
An activation function is an element-wise mathematical function and plays a crucial role in deep neural networks (DNN). Many novel and sophisticated activation functions have been…
cs.CL2022
Transkimmer: Transformer Learns to Layer-wise Skim
Yue Guan, Zhengyi Li, Jingwen Leng +2
Transformer architecture has become the de-facto model for many machine learning tasks from natural language processing and computer vision. As such, improving its computational ef…
cs.LG2022★ 19 cited
SQuant: On-the-Fly Data-Free Quantization via Diagonal Hessian Approximation
Cong Guo, Yuxian Qiu, Jingwen Leng +6
Quantization of deep neural networks (DNN) has been proven effective for compressing and accelerating DNN models. Data-free quantization (DFQ) is a promising approach without the o…