1 paper
Abrar Majeedi, Zhiyuan Ruan, Ziyi Zhao +3
Multimodal large language models (MLLMs) have achieved impressive performance on visual perception and reasoning tasks with RGB imagery, yet they remain fragile under common degrad…