3 papers
cs.CV2026
GT-PCQA: Geometry-Texture Decoupled Point Cloud Quality Assessment with MLLM
Guohua Zhang, Jian Jin, Meiqin Liu +3
With the rapid advancement of Multi-modal Large Language Models (MLLMs), MLLM-based Image Quality Assessment (IQA) methods have shown promising generalization. However, directly ex…
cs.CV2025
Explanatory Instructions: Towards Unified Vision Tasks Understanding and Zero-shot Generalization
Yang Shen, Xiu-Shen Wei, Yifan Sun +6
Computer Vision (CV) has yet to fully achieve the zero-shot task generalization observed in Natural Language Processing (NLP), despite following many of the milestones established…
cs.CV2025
Twofold Debiasing Enhances Fine-Grained Learning with Coarse Labels
Xin-yang Zhao, Jian Jin, Yang-yang Li +1
The Coarse-to-Fine Few-Shot (C2FS) task is designed to train models using only coarse labels, then leverages a limited number of subclass samples to achieve fine-grained recognitio…