1 paper
Jingxuan Li, Yuning Yang, Shengqi Yang +2
The recent progress in Vision-Language Models (VLMs) has broadened the scope of multimodal applications. However, evaluations often remain limited to functional tasks, neglecting a…