1 paper
Zhiqiu Lin, Xinyue Chen, Deepak Pathak +2
Vision-language models (VLMs) are impactful in part because they can be applied to a variety of visual understanding tasks in a zero-shot fashion, without any fine-tuning. We study…