1 paper · 1 filter
Fengbin Zhu, Ziyang Liu, Xiang Yao Ng +6
Large Vision-Language Models (LVLMs) have achieved remarkable performance in many vision-language tasks, yet their capabilities in fine-grained visual understanding remain insuffic…