3 citations · 3 across the 3 of their papers we have counts for
1 paper · 1 filter
Yunze Man, De-An Huang, Guilin Liu +6
Recent advances in multimodal large language models (MLLMs) have demonstrated remarkable capabilities in vision-language tasks, yet they often struggle with vision-centric scenario…