1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Xinyao Yu, Hao Sun, Ziwei Niu +4
In recent years, large-scale pre-trained multimodal models (LMM) generally emerge to integrate the vision and language modalities, achieving considerable success in various natural…