1 paper
Aaditya Baranwal, Vishal Yadav, Abhishek Rajora
While Vision-Language Models (VLMs) demonstrate remarkable zero-shot recognition capabilities across a diverse spectrum of multimodal tasks, it yet remains an open question whether…