2 papers
cs.CV2025
IPTQ-ViT: Post-Training Quantization of Non-linear Functions for Integer-only Vision Transformers
Gihwan Kim, Jemin Lee, Hyungshin Kim
Previous Quantization-Aware Training (QAT) methods for vision transformers rely on expensive retraining to recover accuracy loss in non-linear layer quantization, limiting their us…
cs.CV2025
Mixed Non-linear Quantization for Vision Transformers
Gihwan Kim, Jemin Lee, Sihyeong Park +2
The majority of quantization methods have been proposed to reduce the model size of Vision Transformers, yet most of them have overlooked the quantization of non-linear operations.…