2 papers
cs.CV2025
Post-Training Quantization via Residual Truncation and Zero Suppression for Diffusion Models
Donghoon Kim, Dongyoung Lee, Ik Joon Chang +1
Diffusion models achieve high-quality image generation but face deployment challenges due to their high computational requirements. Although 8-bit outlier-aware post-training quant…
cs.LG2025
Qrazor: Reliable and Effortless 4-bit LLM Quantization by Significant Data Razoring
Dongyoung Lee, Seungkyu Choi, Ik Joon Chang
Large-scale language models (LLMs) excel in language processing tasks but face deployment challenges due to high memory and computational demands. While low-bit quantization, such…