Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
Bridging Compressed Image Latents and Multimodal Large Language Models
Chia-Hao Kao, Cheng Chien, Yu-Jen Tseng +5
This paper presents the first-ever study of adapting compressed image latents to suit the needs of downstream vision tasks that adopt Multimodal Large Language Models (MLLMs). MLLM…
cs.CV2023
Transformer-based Image Compression with Variable Image Quality Objectives
Chia-Hao Kao, Yi-Hsin Chen, Cheng Chien +2
This paper presents a Transformer-based image compression system that allows for a variable image quality objective according to the user's preference. Optimizing a learned codec f…