1 paper
Peiyu Yang, Paul Montague, Feng Liu +4
Vision-language models (VLMs) are now widely used in downstream tasks. However, real-world applications often expose VLMs to distribution shifts induced by semantic variation (e.g.…