activity
20242026
collaborators

7 papers

cs.CV2026

CSGaussian: Progressive Rate-Distortion Compression and Segmentation for 3D Gaussian Splatting

Yu-Jen Tseng, Chia-Hao Kao, Jing-Zhong Chen +4

We present the first unified framework for rate-distortion-optimized compression and segmentation of 3D Gaussian Splatting (3DGS). While 3DGS has proven effective for both real-tim…

eess.IV2025

End-to-End Semantic Preservation in Text-Aware Image Compression Systems

Stefano Della Fiore, Alessandro Gnutti, Marco Dalai +2

Traditional image compression methods aim to reconstruct images for human perception, prioritizing visual fidelity over task relevance. In contrast, Coding for Machines focuses on…

eess.IV2025

MH-LVC: Multi-Hypothesis Temporal Prediction for Learned Conditional Residual Video Coding

Huu-Tai Phung, Zong-Lin Gao, Yi-Chen Yao +5

This work, termed MH-LVC, presents a multi-hypothesis temporal prediction scheme that employs long- and short-term reference frames in a conditional residual video coding framework…

eess.IV2025

Is JPEG AI going to change image forensics?

Edoardo Daniele Cannas, Sara Mandelli, Nataša Popović +4

In this paper, we investigate the counter-forensic effects of the new JPEG AI standard based on neural image compression, focusing on two critical areas: deepfake image detection a…

cs.CV2025

Bridging Compressed Image Latents and Multimodal Large Language Models

Chia-Hao Kao, Cheng Chien, Yu-Jen Tseng +5

This paper presents the first-ever study of adapting compressed image latents to suit the needs of downstream vision tasks that adopt Multimodal Large Language Models (MLLMs). MLLM…

eess.IV2024

Learning Optimal Linear Block Transform by Rate Distortion Minimization

Alessandro Gnutti, Chia-Hao Kao, Wen-Hsiao Peng +1

Linear block transform coding remains a fundamental component of image and video compression. Although the Discrete Cosine Transform (DCT) is widely employed in all current compres…