12 citations · 12 across the 1 of their papers we have counts for
1 paper
Wenbo Hu, Yifan Xu, Yi Li +3
Vision Language Models (VLMs), which extend Large Language Models (LLM) by incorporating visual understanding capability, have demonstrated significant advancements in addressing o…