1 paper
Mehmet Aktukmak, Daniel Huang, Ke Ding
Quantization is essential for reducing the computational cost and memory usage of deep neural networks, enabling efficient inference on low-precision hardware. Despite the growing…