1 paper
Md. Atabuzzaman, Andrew Zhang, Chris Thomas
Large Vision-Language Models (LVLMs) have demonstrated impressive performance on vision-language reasoning tasks. However, their potential for zero-shot fine-grained image classifi…