2 papers
cs.CV2026
Diagnosing Corruption-Induced Reliability Failures in Vision-Language Models
Xiangjie Sui, Songyang Li, Hanwei Zhu +3
Visual corruptions can change vision--language model (VLM) behavior in ways that top-1 accuracy does not capture. A model may keep the same answer while losing distributional suppo…
cs.CV2025
Mitigating Perception Bias: A Training-Free Approach to Enhance LMM for Image Quality Assessment
Baoliang Chen, Siyi Pan, Dongxu Wu +4
Despite the impressive performance of large multimodal models (LMMs) in high-level visual tasks, their capacity for image quality assessment (IQA) remains limited. One main reason…