1 paper · 1 filter
Zhenxiang Lin, Maryam Haghighat, Will Browne +1
Vision-language models (VLMs), such as CLIP, have gained popularity for their strong open vocabulary classification performance, but they are prone to assigning high confidence sco…