1 paper
Heqing Zou, Tianze Luo, Guiyang Xie +8
The integration of Large Language Models (LLMs) with visual encoders has recently shown promising performance in visual understanding tasks, leveraging their inherent capability to…