1 citations · 1 across the 1 of their papers we have counts for
1 paper
Zhelun Shi, Zhipin Wang, Hongxing Fan +4
Multimodal Large Language Models (MLLMs) have shown impressive abilities in interacting with visual content with myriad potential downstream tasks. However, even though a list of b…