27 citations · 27 across the 1 of their papers we have counts for
1 paper
Can Cui, Yunsheng Ma, Xu Cao +18
With the emergence of Large Language Models (LLMs) and Vision Foundation Models (VFMs), multimodal AI systems benefiting from large models have the potential to equally perceive th…