159 citations · 254 across the 18 of their papers we have counts for
19 papers
Enhancing Image Quality Assessment Ability of LMMs via Retrieval-Augmented Generation
Kang Fu, Huiyu Duan, Zicheng Zhang +5
Large Multimodal Models (LMMs) have recently shown remarkable promise in low-level visual perception tasks, particularly in Image Quality Assessment (IQA), demonstrating strong zer…
Robust Mesh Saliency Ground Truth Acquisition in VR via View Cone Sampling and Manifold Diffusion
Guoquan Zheng, Jie Hao, Huiyu Duan +7
As the complexity of 3D digital content grows exponentially, understanding human visual attention is critical for optimizing rendering and processing resources. Therefore, reliable…
VQualA 2025 Challenge on Visual Quality Comparison for Large Multimodal Models: Methods and Results
Hanwei Zhu, Haoning Wu, Zicheng Zhang +26
This paper presents a summary of the VQualA 2025 Challenge on Visual Quality Comparison for Large Multimodal Models (LMMs), hosted as part of the ICCV 2025 Workshop on Visual Quali…
Can Large Models Fool the Eye? A New Turing Test for Biological Animation
Zijian Chen, Lirong Deng, Zhengyu Chen +5
Evaluating the abilities of large models and manifesting their gaps are challenging. Current benchmarks adopt either ground-truth-based score-form evaluation on static datasets or…
AGHI-QA: A Subjective-Aligned Dataset and Metric for AI-Generated Human Images
Yunhao Li, Sijing Wu, Wei Sun +6
The rapid development of text-to-image (T2I) generation approaches has attracted extensive interest in evaluating the quality of generated images, leading to the development of var…
Omni: Unifying Omnidirectional Image Generation and Editing in an Omni Model
Liu Yang, Huiyu Duan, Yucheng Zhu +7
omnidirectional images (ODIs) have gained considerable attention recently, and are widely used in various virtual reality (VR) and augmented reality (AR) applications…