1 paper · 1 filter
Jihwan Bang, Sumyeong Ahn, Jae-Gil Lee
Pre-trained Vision Language Models (VLMs) have demonstrated notable progress in various zero-shot tasks, such as classification and retrieval. Despite their performance, because im…