1 citations · 1 across the 2 of their papers we have counts for
5 papers
Diagnosing Corruption-Induced Reliability Failures in Vision-Language Models
Xiangjie Sui, Songyang Li, Hanwei Zhu +3
Visual corruptions can change vision--language model (VLM) behavior in ways that top-1 accuracy does not capture. A model may keep the same answer while losing distributional suppo…
Mitigating Perception Bias: A Training-Free Approach to Enhance LMM for Image Quality Assessment
Baoliang Chen, Siyi Pan, Dongxu Wu +4
Despite the impressive performance of large multimodal models (LMMs) in high-level visual tasks, their capacity for image quality assessment (IQA) remains limited. One main reason…
2AFC Prompting of Large Multimodal Models for Image Quality Assessment
Hanwei Zhu, Xiangjie Sui, Baoliang Chen +4
While abundant research has been conducted on improving high-level visual understanding and reasoning capabilities of large multimodal models~(LMMs), their visual quality assessmen…
Perceptual Quality Assessment of 360 Images Based on Generative Scanpath Representation
Xiangjie Sui, Hanwei Zhu, Xuelin Liu +3
Despite substantial efforts dedicated to the design of heuristic models for omnidirectional (i.e., 360) image quality assessment (OIQA), a conspicuous gap remains due to th…
Perceptual Quality Assessment of Omnidirectional Images as Moving Camera Videos
Xiangjie Sui, Kede Ma, Yiru Yao +1
Omnidirectional images (also referred to as static 360° panoramas) impose viewing conditions much different from those of regular 2D images. How do humans perceive image distortion…