1 paper
Benjamin Schneider, Florian Kerschbaum, Wenhu Chen
Visual embedding models excel at zero-shot tasks like visual retrieval and classification. However, these models cannot be used for tasks that contain ambiguity or require user ins…