1 paper
Chia-Hao Kao, Cheng Chien, Yu-Jen Tseng +5
This paper presents the first-ever study of adapting compressed image latents to suit the needs of downstream vision tasks that adopt Multimodal Large Language Models (MLLMs). MLLM…