1 paper · 1 filter
Yifan Du, Hangyu Guo, Kun Zhou +6
Visual instruction tuning is crucial for enhancing the zero-shot generalization capability of Multi-modal Large Language Models (MLLMs). In this paper, we aim to investigate a fund…