1 paper
Yunqi Hong, Sohyun An, Andrew Bai +2
Despite Multimodal Large Language Models (MLLMs) showing promising results on general zero-shot image classification tasks, fine-grained image classification remains challenging. I…