1 paper
Eric Yang Yu, Christopher Liao, Sathvik Ravi +2
Recent advances in vision-language models have combined contrastive approaches with generative methods to achieve state-of-the-art (SOTA) on downstream inference tasks like zero-sh…