12 citations · 13 across the 12 of their papers we have counts for
1 paper · 2 filters
King Zhu, Qianbo Zang, Shian Jia +18
Multimodal Large Language Models (MLLMs) are evaluated on various benchmarks, such as image captioning, visual question answering, and reasoning. However, many of these benchmarks…