Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
QGen: On the Ability to Generalize in Quantization Aware Training
MohammadHossein AskariHemmat, Ahmadreza Jeddi, Reyhane Askari Hemmat +6
Quantization lowers memory usage, computational requirements, and latency by utilizing fewer bits to represent model weights and activations. In this work, we investigate the gener…
cs.LG2023
DeepliteRT: Computer Vision at the Edge
Saad Ashfaq, Alexander Hoffman, Saptarshi Mitra +3
The proliferation of edge devices has unlocked unprecedented opportunities for deep learning model deployment in computer vision applications. However, these complex models require…