2 citations · 2 across the 1 of their papers we have counts for
1 paper
Sicong Leng, Hang Zhang, Guanzheng Chen +4
Large Vision-Language Models (LVLMs) have advanced considerably, intertwining visual recognition and language understanding to generate content that is not only coherent but also c…