1 paper · 1 filter
Chenjiao Tan, Qian Cao, Yiwei Li +15
The advent of large language models (LLMs) has heightened interest in their potential for multimodal applications that integrate language and vision. This paper explores the capabi…