19 citations · 19 across the 1 of their papers we have counts for
2 papers
cs.LG2022★ 19 cited
SQuant: On-the-Fly Data-Free Quantization via Diagonal Hessian Approximation
Cong Guo, Yuxian Qiu, Jingwen Leng +6
Quantization of deep neural networks (DNN) has been proven effective for compressing and accelerating DNN models. Data-free quantization (DFQ) is a promising approach without the o…
cs.LG2020
OpEvo: An Evolutionary Method for Tensor Operator Optimization
Xiaotian Gao, Cui Wei, Lintao Zhang +1
Training and inference efficiency of deep neural networks highly rely on the performance of tensor operators on hardware platforms. Manually optimizing tensor operators has limitat…