1 paper
Srija Mukhopadhyay, Abhishek Rajgaria, Prerana Khatiwada +2
Vision-language models (VLMs) excel at tasks requiring joint understanding of visual and linguistic information. A particularly promising yet under-explored application for these m…