1 paper · 1 filter
Jiaqi Zhu, Shaofeng Cai, Fang Deng +2
Large vision-language models (LVLMs) are markedly proficient in deriving visual representations guided by natural language. Recent explorations have utilized LVLMs to tackle zero-s…