1 paper
Peng Gao, Yujian Lee, Xiaofeng Zhang +2
Large Vision-Language Models (LVLMs) have achieved impressive performance across a wide range of multimodal tasks. However, they still face critical challenges in modeling long-ran…