1 citations · 1 across the 9 of their papers we have counts for
1 paper · 1 filter
Yichi Zhang, Zhuo Chen, Lingbing Guo +2
Understanding and reasoning with abstractive information from the visual modality presents significant challenges for current multi-modal large language models (MLLMs). Among the v…