1 paper · 1 filter
Wenqi Shao, Meng Lei, Yutao Hu +8
Recent advancements in Large Vision-Language Models (LVLMs) have demonstrated significant progress in tackling complex multimodal tasks. Among these cutting-edge developments, Goog…