1 paper
Daiqing Qi, Handong Zhao, Zijun Wei +1
Despite recent advances in the general visual instruction-following ability of Multimodal Large Language Models (MLLMs), they still struggle with critical problems when required to…