1 paper
Yatai Ji, Shilong Zhang, Jie Wu +6
The rapid advancement of Large Vision-Language models (LVLMs) has demonstrated a spectrum of emergent capabilities. Nevertheless, current models only focus on the visual content of…