1 paper
Zeqing Wang, Wentao Wan, Qiqing Lao +5
Recently, to comprehensively improve Vision Language Models (VLMs) for Visual Question Answering (VQA), several methods have been proposed to further reinforce the inference capabi…