1 paper
Zijie Zhou, Dandan Zhu, Hangxiangpan Wang +3
Large Vision-Language Models (LVLMs) have demonstrated impressive performance on multimodal tasks through scaled architectures and extensive training. Recent studies introduce Mixt…