1 paper · 1 filter
MohammadHossein AskariHemmat, Ahmadreza Jeddi, Reyhane Askari Hemmat +6
Quantization lowers memory usage, computational requirements, and latency by utilizing fewer bits to represent model weights and activations. In this work, we investigate the gener…