Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
LampQ: Towards Accurate Layer-wise Mixed Precision Quantization for Vision Transformers
Minjun Kim, Jaeri Lee, Jongjin Kim +3
How can we accurately quantize a pre-trained Vision Transformer model? Quantization algorithms compress Vision Transformers (ViTs) into low-bit formats, reducing memory and computa…
cs.CV2025
Zero-shot Quantization: A Comprehensive Survey
Minjun Kim, Jaehyeon Choi, Jongkeun Lee +2
Network quantization has proven to be a powerful approach to reduce the memory and computational demands of deep learning models for deployment on resource-constrained devices. How…