7 papers
CSGaussian: Progressive Rate-Distortion Compression and Segmentation for 3D Gaussian Splatting
Yu-Jen Tseng, Chia-Hao Kao, Jing-Zhong Chen +4
We present the first unified framework for rate-distortion-optimized compression and segmentation of 3D Gaussian Splatting (3DGS). While 3DGS has proven effective for both real-tim…
End-to-End Semantic Preservation in Text-Aware Image Compression Systems
Stefano Della Fiore, Alessandro Gnutti, Marco Dalai +2
Traditional image compression methods aim to reconstruct images for human perception, prioritizing visual fidelity over task relevance. In contrast, Coding for Machines focuses on…
MH-LVC: Multi-Hypothesis Temporal Prediction for Learned Conditional Residual Video Coding
Huu-Tai Phung, Zong-Lin Gao, Yi-Chen Yao +5
This work, termed MH-LVC, presents a multi-hypothesis temporal prediction scheme that employs long- and short-term reference frames in a conditional residual video coding framework…
Is JPEG AI going to change image forensics?
Edoardo Daniele Cannas, Sara Mandelli, NataÅ¡a PopoviÄ +4
In this paper, we investigate the counter-forensic effects of the new JPEG AI standard based on neural image compression, focusing on two critical areas: deepfake image detection a…
Bridging Compressed Image Latents and Multimodal Large Language Models
Chia-Hao Kao, Cheng Chien, Yu-Jen Tseng +5
This paper presents the first-ever study of adapting compressed image latents to suit the needs of downstream vision tasks that adopt Multimodal Large Language Models (MLLMs). MLLM…
Learning Optimal Linear Block Transform by Rate Distortion Minimization
Alessandro Gnutti, Chia-Hao Kao, Wen-Hsiao Peng +1
Linear block transform coding remains a fundamental component of image and video compression. Although the Discrete Cosine Transform (DCT) is widely employed in all current compres…