1 paper
Qian Chen, Xianyin Zhang, Yanzhi Liu +3
The emergence of Large Vision-Language Models (LVLMs) has substantially expanded model capabilities beyond text-only understanding, enabling unified inference across both visual an…