15 citations · 20 across the 8 of their papers we have counts for
1 paper · 1 filter
Junda Wu, Zhehao Zhang, Yu Xia +12
Multimodal large language models (MLLMs) equip pre-trained large-language models (LLMs) with visual capabilities. While textual prompting in LLMs has been widely studied, visual pr…