1 paper
Yuliang Liu, Zhang Li, Mingxin Huang +7
Large models have recently played a dominant role in natural language processing and multimodal vision-language learning. However, their effectiveness in text-related visual tasks…