1 paper
Yuehao Yin, Huiyan Qi, Bin Zhu +3
Large Multi-modal Models (LMMs) have made impressive progress in many vision-language tasks. Nevertheless, the performance of general LMMs in specific domains is still far from sat…