1 paper
Zhaopeng Gu, Bingke Zhu, Guibo Zhu +3
Large Vision-Language Models (LVLMs) such as MiniGPT-4 and LLaVA have demonstrated the capability of understanding images and achieved remarkable performance in various visual task…