1 citations · 1 across the 1 of their papers we have counts for
1 paper
Mubashir Noman, Noor Ahsan, Muzammal Naseer +4
Large multimodal models (LMMs) have shown encouraging performance in the natural image domain using visual instruction tuning. However, these LMMs struggle to describe the content…