27 citations · 27 across the 1 of their papers we have counts for
1 paper
Feng Li, Renrui Zhang, Hao Zhang +5
Visual instruction tuning has made considerable strides in enhancing the capabilities of Large Multimodal Models (LMMs). However, existing open LMMs largely focus on single-image t…