1 paper
Kejia Zhang, Keda Tao, Jiasheng Tang +1
Large vision-language models (LVMs) extend large language models (LLMs) with visual perception capabilities, enabling them to process and interpret visual information. A major chal…