1 paper · 1 filter
A K Nirala, A Joshi, C Hegde +1
A key benefit of deep vision-language models such as CLIP is that they enable zero-shot open vocabulary classification; the user has the ability to define novel class labels via na…